Comprehensive benchmark for RAG
☆301Jun 14, 2025Updated last year
Alternatives and similar repositories for CRAG
Users that are interested in CRAG are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆53Aug 14, 2024Updated 2 years ago
- ☆236Apr 2, 2025Updated last year
- Repository for "MultiHop-RAG: A Dataset for Evaluating Retrieval-Augmented Generation Across Documents" (COLM 2024)☆462Jul 17, 2026Updated last month
- ☆60Jan 19, 2025Updated last year
- RECOMP: Improving Retrieval-Augmented LMs with Compression and Selective Augmentation.☆149Jan 6, 2026Updated 7 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Official repository for RAG-Gym☆127Jul 14, 2026Updated last month
- ☆65Jul 10, 2025Updated last year
- Multi-Turn RAG Benchmark☆151Jul 14, 2026Updated last month
- 🔍 Search-o1: Agentic Search-Enhanced Large Reasoning Models [EMNLP 2025]☆1,241Nov 17, 2025Updated 9 months ago
- Corrective Retrieval Augmented Generation☆468Oct 8, 2024Updated last year
- ☆372May 17, 2024Updated 2 years ago
- Benchmarking library for RAG☆276Jul 14, 2026Updated last month
- ECIR 2024: Sparse lexical representation for image-text retrieval☆13Jul 8, 2024Updated 2 years ago
- ☆11Sep 10, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implementation for "Law of the Weakest Link: Cross capabilities of Large Language Models"☆43Oct 1, 2024Updated last year
- meta-comprehensive-rag-benchmark-kdd-cup-2024 phase1 task1 rank3☆21Jun 21, 2024Updated 2 years ago
- CRUD-RAG: A Comprehensive Chinese Benchmark for Retrieval-Augmented Generation of Large Language Models☆402May 20, 2025Updated last year
- ☆14Apr 16, 2024Updated 2 years ago
- ☆21Apr 17, 2023Updated 3 years ago
- [ACL 2025] AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark☆167Mar 29, 2026Updated 4 months ago
- Code and Data for "FaithfulRAG: Fact-Level Conflict Modeling for Context-Faithful Retrieval-Augmented Generation" (ACL25)☆39Oct 26, 2025Updated 9 months ago
- [ICLR 2025] BRIGHT: A Realistic and Challenging Benchmark for Reasoning-Intensive Retrieval☆210Sep 13, 2025Updated 11 months ago
- RAGChecker: A Fine-grained Framework For Diagnosing RAG☆1,107Dec 13, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Automated Evaluation of RAG Systems☆732Mar 28, 2025Updated last year
- [SIGIR 2024] The official repo for paper "Planning Ahead in Generative Retrieval: Guiding Autoregressive Generation through Simultaneous …☆32Apr 24, 2024Updated 2 years ago
- code for paper "Discerning and Resolving Knowledge Conflicts through Adaptive Decoding with Contextual Information-Entropy Constraint"☆12Sep 29, 2024Updated last year
- Official source code repository for paper BubbleRAG.☆17Jun 1, 2026Updated 2 months ago
- Supercharge Your LLM Application Evaluations 🚀☆15,342Feb 24, 2026Updated 5 months ago
- [EMNLP 2024 (Oral)] Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA☆155Dec 22, 2025Updated 7 months ago
- Zero-Shot Learning in Named Entity Recognition with Common Sense Knowledge☆17Nov 16, 2021Updated 4 years ago
- LongBench v2 and LongBench (ACL 25'&24')☆1,223Jan 15, 2025Updated last year
- The official repository of the paper "Do Reasoning Models Enhance Embedding Models?"☆31Apr 17, 2026Updated 4 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Scaling Deep Research via Reinforcement Learning in Real-world Environments.☆795May 10, 2026Updated 3 months ago
- Dense X Retrieval: What Retrieval Granularity Should We Use?☆171Jan 8, 2024Updated 2 years ago
- Glottolog data as CLDF StructureDataset☆17Mar 2, 2026Updated 5 months ago
- Official repository for "Scaling Retrieval-Based Langauge Models with a Trillion-Token Datastore".☆226Dec 16, 2025Updated 8 months ago
- Benchmark baseline for retrieval qa applications☆121Apr 14, 2024Updated 2 years ago
- ☆202Jun 2, 2025Updated last year
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,302Nov 13, 2025Updated 9 months ago