Dingo: A Comprehensive AI Data, Model and Application Quality Evaluation Tool
☆736Aug 10, 2026Updated this week
Alternatives and similar repositories for dingo
Users that are interested in dingo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Web Structured Data Extraction Agent☆16Mar 10, 2026Updated 5 months ago
- A general-purpose API load testing platform that supports LLM services and business HTTP interfaces, enabling one-click performance testi…☆201Jul 10, 2026Updated last month
- WebMainBench is a high-precision benchmark for evaluating web main content extraction.☆20Jun 13, 2026Updated last month
- Autonomous web browser agent that audits performance, functionality & UX for engineers and vibe-coding creators. 网站自主评估测试 Agent,支持 GUI/CL…☆225Jul 2, 2026Updated last month
- MinerU-HTML: An SLM-powered HTML main content extractor that outputs clean HTML bodies. Perfect for Deep Research Agents, RAG application…☆282Mar 27, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The Open-Source Data Annotation Platform☆1,271Jul 2, 2026Updated last month
- [ACL 2025 Best Theme Paper] This is the official implementation for the paper: "Meta-rater: A Multi-dimensional Data Selection Method for…☆196Aug 29, 2025Updated 11 months ago
- Open-source multimodal data annotation platform with AI auto-annotation support.☆1,651Jul 28, 2026Updated 2 weeks ago
- Data processing for and with foundation models! 🍎 🍋 🌽 ➡️ ➡️🍸 🍹 🍷☆6,856Updated this week
- Tools for OpenDataArena: Fair, Open, and Transparent Arena for Data☆145Mar 15, 2026Updated 4 months ago
- WanJuan3.0(“万卷·丝路”)一个作为综合性的纯文本语料库,采集了多个国家地区的网络公开信息、文献、专利等资料,数据总规模超1.2TB,Token总数超过300B,处于国际领先水平,首期开源的语料库主要由泰语、俄语、阿拉伯语、韩语和越南语5个子集构成,每个子集的数据…☆47Feb 13, 2025Updated last year
- [CVPR 2025] A Comprehensive Benchmark for Document Parsing and Evaluation☆1,962Jul 27, 2026Updated 2 weeks ago
- Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL…☆15,102Updated this week
- SDK of OpenDataLab - https://opendatalab.org.cn☆60Jul 31, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ACL 2025] An official pytorch implement of the paper: Condor: Enhance LLM Alignment with Knowledge-Driven Data Synthesis and Refinement☆40May 28, 2025Updated last year
- A powerful tool for creating datasets for LLM fine-tuning 、RAG and Eval☆14,768May 1, 2026Updated 3 months ago
- OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, …☆7,288Aug 4, 2026Updated last week
- Data annotation component library --provided as NPM packages☆160Jul 28, 2026Updated 2 weeks ago
- A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.☆3,219Updated this week
- Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.☆77,281Updated this week
- Easy Data Preparation with latest LLMs-based Operators and Pipelines.☆7,259Updated this week
- Our code for ICLR'25 paper "DataMan: Data Manager for Pre-training Large Language Models".☆130Feb 7, 2026Updated 6 months ago
- Official Repository of "LLM × DATA" Survey Paper☆816Jun 15, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆30Apr 15, 2026Updated 3 months ago
- 基于 MinerU 的智能论文阅读助手,提供 PDF 文档解析、OCR 识别、表格提取等功能。☆19Dec 2, 2025Updated 8 months ago
- ☆25Nov 7, 2022Updated 3 years ago
- Data browser based on s3. 一个基于 S3 的数据(json / jsonl / parquet / html / md等)可视化工具。👇 Try online.☆90Apr 14, 2026Updated 3 months ago
- An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models☆3,356Updated this week
- Data Set Description Language Specification (新一代人工智能数据集描述语言DSDL)☆46May 29, 2024Updated 2 years ago
- Agent framework and applications built upon Qwen>=3.0, featuring Function Calling, MCP, Code Interpreter, RAG, Chrome extension, etc.☆16,943Mar 4, 2026Updated 5 months ago
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆22,900Updated this week
- EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL☆5,107Jul 30, 2026Updated last week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- LMDeploy is a toolkit for compressing, deploying, and serving LLMs.☆7,998Updated this week
- A lightweight framework for building LLM-based agents☆2,274Aug 3, 2026Updated last week
- Tongyi Deep Research, the Leading Open-source Deep Research Agent☆19,807Feb 27, 2026Updated 5 months ago
- Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)☆73,967Updated this week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,653Updated this week
- Ray-powered accelerator for MinerU, turning PDF → Markdown into a scalable, cluster-ready data infrastructure. 基于 Ray 的 MinerU 加速层,将 PDF …☆68Apr 20, 2026Updated 3 months ago
- MPB (Miner-PDF-Benchmark) is an end-to-end PDF document comprehension evaluation suite designed for large-scale model data scenarios.☆24Dec 11, 2024Updated last year