CRUD-RAG: A Comprehensive Chinese Benchmark for Retrieval-Augmented Generation of Large Language Models
☆400May 20, 2025Updated last year
Alternatives and similar repositories for CRUD_RAG
Users that are interested in CRUD_RAG are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PGRAG☆53Jul 16, 2024Updated 2 years ago
- ☆371May 17, 2024Updated 2 years ago
- [ACL 2024 Main] NewsBench: A Systematic Evaluation Framework for Assessing Editorial Capabilities of Large Language Models in Chinese Jou…☆34Jun 25, 2024Updated 2 years ago
- [ACL 2024]Controlled Text Generation for Large Language Model with Dynamic Attribute Graphs☆40Sep 24, 2024Updated last year
- 中文原生检索增强生成测评基准☆132Apr 18, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆23Jun 10, 2025Updated last year
- ☆237Apr 2, 2025Updated last year
- [ACL 2024] User-friendly evaluation framework: Eval Suite & Benchmarks: UHGEval, HaluEval, HalluQA, etc.☆181Jun 7, 2025Updated last year
- ☆61Mar 11, 2025Updated last year
- Automated Evaluation of RAG Systems☆730Mar 28, 2025Updated last year
- Repository for "MultiHop-RAG: A Dataset for Evaluating Retrieval-Augmented Generation Across Documents" (COLM 2024)☆456Jul 17, 2026Updated last week
- ⚡FlashRAG: A Python Toolkit for Efficient RAG Research (WWW2025 Resource)☆3,535Jul 19, 2026Updated last week
- xVerify: Efficient Answer Verifier for Reasoning Model Evaluations☆149Nov 13, 2025Updated 8 months ago
- [ACL 2025] AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark☆167Mar 29, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Grimoire is All You Need for Enhancing Large Language Models☆120Feb 29, 2024Updated 2 years ago
- Evaluation tools for Retrieval-augmented Generation (RAG) methods.☆171Nov 18, 2024Updated last year
- Retrieval and Retrieval-augmented LLMs☆11,990Apr 22, 2026Updated 3 months ago
- Corrective Retrieval Augmented Generation☆467Oct 8, 2024Updated last year
- ☆11May 17, 2024Updated 2 years ago
- [ICLR 2025] xFinder: Large Language Models as Automated Evaluators for Reliable Evaluation☆179Nov 14, 2025Updated 8 months ago
- Codes for our paper "RQ-RAG: Learning to Refine Queries for Retrieval Augmented Generation"☆212Aug 16, 2024Updated last year
- Fast Memorization of Prompt Improves Context Awareness of Large Language Models (Findings of EMNLP 2024)☆22Oct 22, 2024Updated last year
- Supercharge Your LLM Application Evaluations 🚀☆15,016Feb 24, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- HaluMem is the first operation level hallucination evaluation benchmark tailored to agent memory systems.☆148Apr 30, 2026Updated 2 months ago
- This includes the original implementation of SELF-RAG: Learning to Retrieve, Generate and Critique through self-reflection by Akari Asai,…☆2,413May 25, 2024Updated 2 years ago
- ☆2,140May 8, 2024Updated 2 years ago
- Benchmark baseline for retrieval qa applications☆121Apr 14, 2024Updated 2 years ago
- An awesome repository & A comprehensive survey on interpretability of LLM attention heads.☆412Mar 2, 2025Updated last year
- Collecting awesome papers of RAG for AIGC. We propose a taxonomy of RAG foundations, enhancements, and applications in paper "Retrieval-…☆1,788Aug 20, 2024Updated last year
- ☆46Jan 21, 2025Updated last year
- meta-comprehensive-rag-benchmark-kdd-cup-2024 phase1 task1 rank3☆21Jun 21, 2024Updated 2 years ago
- text embedding☆145Sep 18, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- the resources about the application based on LLM with RAG pattern☆1,641Mar 10, 2026Updated 4 months ago
- Explore concepts like Self-Correct, Self-Refine, Self-Improve, Self-Contradict, Self-Play, and Self-Knowledge, alongside o1-like reasonin…☆173Dec 7, 2024Updated last year
- A curated list of resources dedicated to retrieval-augmented generation (RAG).☆136Oct 31, 2025Updated 8 months ago
- aigc evals☆10Dec 2, 2023Updated 2 years ago
- Comprehensive benchmark for RAG☆297Jun 14, 2025Updated last year
- code for piccolo embedding model from SenseTime☆143May 21, 2024Updated 2 years ago
- ☆986Feb 7, 2025Updated last year