Comprehensive benchmark for RAG
☆297Jun 14, 2025Updated last year
Alternatives and similar repositories for CRAG
Users that are interested in CRAG are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆22Jun 12, 2024Updated 2 years ago
- ☆53Aug 14, 2024Updated last year
- ☆237Apr 2, 2025Updated last year
- Repository for "MultiHop-RAG: A Dataset for Evaluating Retrieval-Augmented Generation Across Documents" (COLM 2024)☆456Jul 17, 2026Updated last week
- ☆60Jan 19, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- RECOMP: Improving Retrieval-Augmented LMs with Compression and Selective Augmentation.☆148Jan 6, 2026Updated 6 months ago
- Official repository for RAG-Gym☆126Jul 14, 2026Updated 2 weeks ago
- ☆64Jul 10, 2025Updated last year
- Multi-Turn RAG Benchmark☆149Jul 14, 2026Updated 2 weeks ago
- 🔍 Search-o1: Agentic Search-Enhanced Large Reasoning Models [EMNLP 2025]☆1,240Nov 17, 2025Updated 8 months ago
- Corrective Retrieval Augmented Generation☆467Oct 8, 2024Updated last year
- ☆371May 17, 2024Updated 2 years ago
- Benchmarking library for RAG☆276Jul 14, 2026Updated 2 weeks ago
- ☆11Sep 10, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implementation for "Law of the Weakest Link: Cross capabilities of Large Language Models"☆43Oct 1, 2024Updated last year
- meta-comprehensive-rag-benchmark-kdd-cup-2024 phase1 task1 rank3☆21Jun 21, 2024Updated 2 years ago
- CRUD-RAG: A Comprehensive Chinese Benchmark for Retrieval-Augmented Generation of Large Language Models☆400May 20, 2025Updated last year
- ☆14Apr 16, 2024Updated 2 years ago
- Confidence Regulation Neurons in Language Models (NeurIPS 2024)☆15Feb 1, 2025Updated last year
- ☆21Apr 17, 2023Updated 3 years ago
- [ACL 2025] AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark☆167Mar 29, 2026Updated 4 months ago
- Code and Data for "FaithfulRAG: Fact-Level Conflict Modeling for Context-Faithful Retrieval-Augmented Generation" (ACL25)☆39Oct 26, 2025Updated 9 months ago
- [ICLR 2025] BRIGHT: A Realistic and Challenging Benchmark for Reasoning-Intensive Retrieval☆210Sep 13, 2025Updated 10 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- RAGChecker: A Fine-grained Framework For Diagnosing RAG☆1,102Dec 13, 2024Updated last year
- Automated Evaluation of RAG Systems☆730Mar 28, 2025Updated last year
- [SIGIR 2024] The official repo for paper "Planning Ahead in Generative Retrieval: Guiding Autoregressive Generation through Simultaneous …☆32Apr 24, 2024Updated 2 years ago
- code for paper "Discerning and Resolving Knowledge Conflicts through Adaptive Decoding with Contextual Information-Entropy Constraint"☆12Sep 29, 2024Updated last year
- Official source code repository for paper BubbleRAG.☆16Jun 1, 2026Updated last month
- Supercharge Your LLM Application Evaluations 🚀☆15,016Feb 24, 2026Updated 5 months ago
- [EMNLP 2024 (Oral)] Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA☆155Dec 22, 2025Updated 7 months ago
- Zero-Shot Learning in Named Entity Recognition with Common Sense Knowledge☆17Nov 16, 2021Updated 4 years ago
- LongBench v2 and LongBench (ACL 25'&24')☆1,215Jan 15, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The official repository of the paper "Do Reasoning Models Enhance Embedding Models?"☆30Apr 17, 2026Updated 3 months ago
- Scaling Deep Research via Reinforcement Learning in Real-world Environments.☆784May 10, 2026Updated 2 months ago
- Dense X Retrieval: What Retrieval Granularity Should We Use?☆171Jan 8, 2024Updated 2 years ago
- Glottolog data as CLDF StructureDataset☆17Mar 2, 2026Updated 4 months ago
- Official repository for "Scaling Retrieval-Based Langauge Models with a Trillion-Token Datastore".☆226Dec 16, 2025Updated 7 months ago
- Benchmark baseline for retrieval qa applications☆121Apr 14, 2024Updated 2 years ago
- ☆201Jun 2, 2025Updated last year