An End-to-End Benchmarking Framework for Retrieval-Augmented Generation Systems
☆31Mar 13, 2026Updated 6 months ago
Alternatives and similar repositories for RAGPerf
Users that are interested in RAGPerf are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Jun 1, 2022Updated 4 years ago
- ☆19Feb 9, 2026Updated 7 months ago
- Prefix-Aware Attention for LLM Decoding☆47May 26, 2026Updated 3 months ago
- An open-source simulator framework for neural processing units☆55Sep 10, 2026Updated last week
- An FPGA-based full-stack in-storage computing system.☆37Nov 6, 2020Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Clue-RAG: Towards Accurate and Cost-Efficient Graph-based RAG via Multi-Partite Graph and Query-Driven Iterative Retrieval☆26Mar 3, 2026Updated 6 months ago
- The code implementation of HyGRAG, accepted by WWW'26.☆15May 31, 2026Updated 3 months ago
- ☆12Aug 1, 2022Updated 4 years ago
- ☆17Jun 15, 2026Updated 3 months ago
- ☆37Apr 10, 2024Updated 2 years ago
- A guide on how to emulate an NVMe SPDM responder device with QEMU and Linux. Additionally, instructions on setting up and testing the (in…☆12Sep 3, 2024Updated 2 years ago
- Open-source Verifiable Data Structures Server implementation☆13Aug 12, 2025Updated last year
- ☆14Apr 24, 2024Updated 2 years ago
- ☆20Mar 11, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆26Jan 10, 2023Updated 3 years ago
- ☆42Updated this week
- ☆10May 12, 2022Updated 4 years ago
- A low-cost, high-performance deep learning training framework that enables efficient 100B-scale model fine-tuning on a commodity server w…☆23Mar 21, 2025Updated last year
- [ASPLOS'26] HILOS: A Cost-Effective Near-Storage Processing Solution for Offline Inference of Long-Context LLMs☆22Jan 18, 2026Updated 8 months ago
- DEDISbench: A disk I/O block-based benchmark for deduplication systems. Unlike other existing benchmarks, written content is generated i…☆14Jul 22, 2021Updated 5 years ago
- ☆14Jul 13, 2025Updated last year
- Agentic RAG Harness for long documents, Tree and Graph based reasoning. Cited answers down to the pixel☆70Apr 15, 2026Updated 5 months ago
- Doc Thinker: All-in-One RAG - document parsing, graph RAG, and evaluation☆32Updated this week
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- multicast learning in network programming course☆10Oct 30, 2020Updated 5 years ago
- Efficient and modular GraphRAG system☆50Jul 7, 2026Updated 2 months ago
- Convert regular expressions to minimized DFAs in AT&T FST format.☆15Sep 12, 2026Updated last week
- AT2PO: Agentic Turn-based Policy Optimization via Tree Search☆22May 21, 2026Updated 3 months ago
- ☆15Mar 19, 2022Updated 4 years ago
- [ACL 2026] WildGraphBench: Benchmarking GraphRAG with Wild-Source Corpora☆18May 11, 2026Updated 4 months ago
- The Gtgraph library from Georgia Tech.☆18Mar 30, 2023Updated 3 years ago
- vLLM fork with Marlin W4A8 SM121 patches + TMA module☆15Mar 23, 2026Updated 5 months ago
- Hardware Accelerators (HwAs) constructed in Vivado HLS☆20Jul 17, 2017Updated 9 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Example of applying CUDA graphs to LLaMA-v2☆11Aug 25, 2023Updated 3 years ago
- Code for the paper "A is for Absorption: Studying Feature Splitting and Absorption in Sparse Autoencoders"☆17Dec 28, 2025Updated 8 months ago
- 基于 BNF 的语法高亮☆18Jul 26, 2026Updated last month
- NVFP4 inference on Blackwell GeForce (RTX 5090/5080/5070 Ti/RTX PRO 6000) — SM120 patches for vLLM + FlashInfer + CUTLASS. 175 tok/s on Q…☆25Apr 27, 2026Updated 4 months ago
- Agent application/benchmark/workload traces should be placed here.☆15Apr 13, 2026Updated 5 months ago
- Data and code for ACL 2026 Paper "Rethinking Reasoning-Intensive Retrieval: Evaluating and Advancing Retrievers in Agentic Search Systems…☆20Apr 30, 2026Updated 4 months ago
- ☆29May 18, 2021Updated 5 years ago