Official repository for the ICLR 2026 Oral Paper🔥 “Q-RAG: Long Context Multi-Step Retrieval via Value-Based Embedder Training”
☆61Sep 4, 2026Updated 2 weeks ago
Alternatives and similar repositories for Q-RAG
Users that are interested in Q-RAG are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- HiPRAG (Hierarchical Process Rewards for Efficient Agentic Retrieval Augmented Generation) is a reinforcement learning method designed fo…☆27Oct 10, 2025Updated 11 months ago
- ☆13Jul 22, 2024Updated 2 years ago
- ☆27Nov 20, 2025Updated 10 months ago
- LoongRL: Reinforcement Learning for Advanced Reasoning over Long Contexts (ICLR 2026 Oral)☆38Feb 20, 2026Updated 6 months ago
- The code for the paper "A Bayesian Approach to Online Planning" published in ICML 2024.☆13Jun 17, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Overflow Prevention Enhances Long-Context Recurrent LLMs (COLM 2025)☆18Jul 8, 2025Updated last year
- Code for paper: Long cOntext aliGnment via efficient preference Optimization☆26Oct 10, 2025Updated 11 months ago
- ☆14Nov 2, 2025Updated 10 months ago
- Tuning the Right Foundation Models is What you Need for Partial Label Learning☆21Nov 2, 2025Updated 10 months ago
- Advancing search on top of AI agents☆35Jun 9, 2026Updated 3 months ago
- Clue-RAG: Towards Accurate and Cost-Efficient Graph-based RAG via Multi-Partite Graph and Query-Driven Iterative Retrieval☆26Mar 3, 2026Updated 6 months ago
- Look Back to Reason Forward: Revisitable Memory for Long-Context LLM Agents☆45Apr 13, 2026Updated 5 months ago
- ☆17Jun 15, 2026Updated 3 months ago
- The code implementation for TTCS: Test-Time Curriculum Synthesis for Self-Evolving.☆51Apr 22, 2026Updated 4 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- something for paper agent☆11Dec 18, 2024Updated last year
- ☆11Aug 20, 2025Updated last year
- (ACL 2025) Divide-Then-Aggregate: An Efficient Tool Learning Method via Parallel Tool Invocation☆12May 21, 2025Updated last year
- 🎓Automatically Update agent Papers Daily using Github Actions (Update Every 12th hours)每日更新agent相关论文(已附带中文摘要翻译)☆37Updated this week
- [ICML 2025] "From Passive to Active Reasoning: Can Large Language Models Ask the Right Questions under Incomplete Information?"☆48Oct 8, 2025Updated 11 months ago
- [COLING 2024] SentiCSE: A Sentiment-aware Contrastive Sentence Embedding Framework with Sentiment-guided Textual Similarity☆13May 8, 2024Updated 2 years ago
- An End-to-End Benchmarking Framework for Retrieval-Augmented Generation Systems☆31Mar 13, 2026Updated 6 months ago
- ☆10Nov 1, 2021Updated 4 years ago
- [ICML'26] Hierarchical Abstract Tree for Cross-Document Retrieval Augmented Generation☆36Aug 28, 2026Updated 3 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Breaking Certifiable Defenses☆17Nov 22, 2022Updated 3 years ago
- ☆16Sep 4, 2024Updated 2 years ago
- [ICLR 2025] This is the code repo for our ICLR’25 paper "RAG-DDR: Optimizing Retrieval-Augmented Generation Using Differentiable Data Rew…☆55Feb 10, 2025Updated last year
- RL Environment and Benchmark pipeline. Code accompanying paper "Neurophysiologically Realistic Environment for Comparing Adaptive Deep Br…☆20Feb 6, 2026Updated 7 months ago
- 由中国政法大学和北京航空航天大学共同设计,基于GLM-9B的法律文书处理和判决预测模型☆29Sep 6, 2024Updated 2 years ago
- ☆10Feb 17, 2024Updated 2 years ago
- ☆18May 11, 2021Updated 5 years ago
- Incremental Mobile User Profiling: Reinforcement Learning with Spatial Knowledge Graph for Modeling Event Streams☆15Jul 25, 2024Updated 2 years ago
- Code repository for the ICML 2026 Oral paper "Characterizing, Evaluating, and Optimizing Complex Reasoning".☆20Jun 21, 2026Updated 2 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This is a fork of SGLang for hip-attention integration. Please refer to hip-attention for detail.☆18Mar 31, 2026Updated 5 months ago
- Code for the paper "A is for Absorption: Studying Feature Splitting and Absorption in Sparse Autoencoders"☆17Dec 28, 2025Updated 8 months ago
- R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning☆82May 25, 2025Updated last year
- pure go for rwkv☆18Dec 31, 2023Updated 2 years ago
- Official source code repository for paper BubbleRAG.☆17Jun 1, 2026Updated 3 months ago
- Go implementation of Recursive Language Models (RLM) - inference-time scaling for arbitrarily long contexts☆20May 12, 2026Updated 4 months ago
- PyTorch Implementation for InMaP☆12Oct 28, 2023Updated 2 years ago