REFRAG-style RAG (compress → sense/select → expand) — Single-file reference implementation
☆235Dec 26, 2025Updated 8 months ago
Alternatives and similar repositories for REFRAG
Users that are interested in REFRAG are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Python Implementation for Rethinking RAG based Decoding☆17Sep 10, 2025Updated last year
- High-performance late-interaction retrieval engine for on-prem AI. ColBERT/ColPali multi-vector search with Rust fused MaxSim, Triton GPU…☆17Jul 6, 2026Updated 2 months ago
- Official code implementation of the paper: QuCo-RAG: Quantifying Uncertainty from the Pre-training Corpus for Dynamic Retrieval-Augmente…☆47Jul 5, 2026Updated 2 months ago
- ☆10Apr 20, 2019Updated 7 years ago
- ☆44Apr 22, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Getting started using pandasai as a Data Copilot☆16Apr 17, 2024Updated 2 years ago
- This project is a versatile and powerful search tool that leverages state-of-the-art natural language processing models to provide releva…☆12Apr 3, 2023Updated 3 years ago
- EMNLP 2024 "Re-reading improves reasoning in large language models". Simply repeating the question to get bidirectional understanding for…☆30Dec 10, 2024Updated last year
- Model souping for LLMs☆75Nov 18, 2025Updated 10 months ago
- Implementation for EACL 2024 paper "Corpus-Steered Query Expansion with Large Language Models"☆13Mar 19, 2024Updated 2 years ago
- a Video Quality Analysis Toolkit☆14May 16, 2025Updated last year
- [AAAI 2026 🔥 Poster] ComoRAG: A Cognitive-Inspired Memory-Organized RAG for Stateful Long Narrative Reasoning☆347Aug 28, 2025Updated last year
- PIKE-RAG: sPecIalized KnowledgE and Rationale Augmented Generation☆2,482Sep 10, 2025Updated last year
- ☆293Nov 6, 2025Updated 10 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆15Nov 19, 2025Updated 10 months ago
- A super fast Graph Database uses GraphBLAS under the hood for its sparse adjacency matrix graph representation. Our goal is to provide th…☆6,180Updated this week
- ☆14Aug 7, 2023Updated 3 years ago
- ☆17Sep 11, 2026Updated last week
- ☆16Feb 6, 2024Updated 2 years ago
- Official Code of Memento: Fine-tuning LLM Agents without Fine-tuning LLMs☆2,580Oct 5, 2025Updated 11 months ago
- English or Chinses GPT2Dialog model from GPT2-chitchat☆12Feb 23, 2020Updated 6 years ago
- ssc-FinLLM-金融大模型☆27Apr 22, 2024Updated 2 years ago
- WhisperMesh is an advanced chatbot that integrates voice and text interactions, delivering personalized responses through LLM models and …☆17Apr 23, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆35Oct 9, 2025Updated 11 months ago
- 📑 PageIndex: Document Index for Vectorless, Reasoning-based RAG☆35,752Updated this week
- The repository of EMNLP 2023 "MixEdit: Revisiting Data Augmentation and Beyond for Grammatical Error Correction"☆12Nov 25, 2023Updated 2 years ago
- The Vision Wormhole: Latent-Space Communication in Heterogeneous Multi-Agent Systems☆29Mar 3, 2026Updated 6 months ago
- Identify which embedding model produced a vector using digit-level tokenization and a tiny transformer☆23Mar 7, 2026Updated 6 months ago
- The official repo of FineSure (ACL-2024)☆36Jul 8, 2024Updated 2 years ago
- ☆48Apr 6, 2025Updated last year
- QRHead: Query-Focused Retrieval Heads Improve Long-Context Reasoning and Re-ranking☆42Jan 20, 2026Updated 8 months ago
- SSRL: Self-Search Reinforcement Learning☆212Aug 20, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆12Apr 20, 2020Updated 6 years ago
- [ICLR 2026] LinearRAG: Linear Graph Retrieval Augmented Generation on Large-scale Corpora☆547Jul 5, 2026Updated 2 months ago
- DeepDive: Advancing Deep Search Agents with Knowledge Graphs and Multi-Turn RL☆349Jun 17, 2026Updated 3 months ago
- Self-Adapting Language Models☆1,860Aug 1, 2025Updated last year
- Span-level grounding verification for RAG, code, and tool-grounded AI outputs.☆611Sep 7, 2026Updated last week
- The absolute trainer to light up AI agents.☆18,375Updated this week
- This repository open-sources our GEC system submitted by THU KELab (sz) in the CCL2023-CLTC Track 1: Multidimensional Chinese Learner Tex…☆15Nov 25, 2023Updated 2 years ago