Comprehensive benchmark for RAG
☆305Jun 14, 2025Updated last year
Alternatives and similar repositories for CRAG
Users that are interested in CRAG are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆24Jun 12, 2024Updated 2 years ago
- ☆53Aug 14, 2024Updated 2 years ago
- ☆239Apr 2, 2025Updated last year
- Repository for "MultiHop-RAG: A Dataset for Evaluating Retrieval-Augmented Generation Across Documents" (COLM 2024)☆473Sep 16, 2026Updated last week
- ☆61Jan 19, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- RECOMP: Improving Retrieval-Augmented LMs with Compression and Selective Augmentation.☆149Jan 6, 2026Updated 8 months ago
- Official repository for RAG-Gym☆128Jul 14, 2026Updated 2 months ago
- ☆67Jul 10, 2025Updated last year
- 🔍 Search-o1: Agentic Search-Enhanced Large Reasoning Models [EMNLP 2025]☆1,248Nov 17, 2025Updated 10 months ago
- Multi-Turn RAG Benchmark☆155Sep 4, 2026Updated 3 weeks ago
- Corrective Retrieval Augmented Generation☆474Oct 8, 2024Updated last year
- ☆378May 17, 2024Updated 2 years ago
- Benchmarking library for RAG☆281Jul 14, 2026Updated 2 months ago
- ECIR 2024: Sparse lexical representation for image-text retrieval☆13Jul 8, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Official implementation for "Law of the Weakest Link: Cross capabilities of Large Language Models"☆43Oct 1, 2024Updated last year
- ☆10Sep 10, 2023Updated 3 years ago
- meta-comprehensive-rag-benchmark-kdd-cup-2024 phase1 task1 rank3☆21Jun 21, 2024Updated 2 years ago
- CRUD-RAG: A Comprehensive Chinese Benchmark for Retrieval-Augmented Generation of Large Language Models☆414May 20, 2025Updated last year
- ☆16Apr 16, 2024Updated 2 years ago
- Confidence Regulation Neurons in Language Models (NeurIPS 2024)☆16Feb 1, 2025Updated last year
- ☆21Apr 17, 2023Updated 3 years ago
- [ACL 2025] AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark☆167Mar 29, 2026Updated 5 months ago
- Code and Data for "FaithfulRAG: Fact-Level Conflict Modeling for Context-Faithful Retrieval-Augmented Generation" (ACL25)☆39Oct 26, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICLR 2025] BRIGHT: A Realistic and Challenging Benchmark for Reasoning-Intensive Retrieval☆214Sep 13, 2025Updated last year
- RAGChecker: A Fine-grained Framework For Diagnosing RAG☆1,124Dec 13, 2024Updated last year
- Automated Evaluation of RAG Systems☆734Mar 28, 2025Updated last year
- [SIGIR 2024] The official repo for paper "Planning Ahead in Generative Retrieval: Guiding Autoregressive Generation through Simultaneous …☆32Apr 24, 2024Updated 2 years ago
- code for paper "Discerning and Resolving Knowledge Conflicts through Adaptive Decoding with Contextual Information-Entropy Constraint"☆12Sep 29, 2024Updated last year
- Official source code repository for paper BubbleRAG.☆17Sep 21, 2026Updated last week
- Supercharge Your LLM Application Evaluations 🚀☆15,859Feb 24, 2026Updated 7 months ago
- [EMNLP 2024 (Oral)] Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA☆154Dec 22, 2025Updated 9 months ago
- Zero-Shot Learning in Named Entity Recognition with Common Sense Knowledge☆17Nov 16, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- LongBench v2 and LongBench (ACL 25'&24')☆1,240Jan 15, 2025Updated last year
- The official repository of the paper "Do Reasoning Models Enhance Embedding Models?"☆32Apr 17, 2026Updated 5 months ago
- Scaling Deep Research via Reinforcement Learning in Real-world Environments.☆800May 10, 2026Updated 4 months ago
- Dense X Retrieval: What Retrieval Granularity Should We Use?☆171Jan 8, 2024Updated 2 years ago
- Glottolog data as CLDF StructureDataset☆17Mar 2, 2026Updated 6 months ago
- Official repository for "Scaling Retrieval-Based Langauge Models with a Trillion-Token Datastore".☆226Dec 16, 2025Updated 9 months ago
- Benchmark baseline for retrieval qa applications☆122Apr 14, 2024Updated 2 years ago