Official repository for the ICLR 2026 Oral Paper🔥 “Q-RAG: Long Context Multi-Step Retrieval via Value-Based Embedder Training”
☆65Sep 4, 2026Updated last month
Alternatives and similar repositories for Q-RAG
Users that are interested in Q-RAG are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- HiPRAG (Hierarchical Process Rewards for Efficient Agentic Retrieval Augmented Generation) is a reinforcement learning method designed fo…☆27Oct 10, 2025Updated last year
- [ICML 2023] Meta-SAGE: Scale Meta-Learning Scheduled Adaptation with Guided Exploration for Mitigating Scale Shift on Combinatorial Optim…☆11Dec 19, 2023Updated 2 years ago
- LoongRL: Reinforcement Learning for Advanced Reasoning over Long Contexts (ICLR 2026 Oral)☆38Feb 20, 2026Updated 7 months ago
- Code for paper: Long cOntext aliGnment via efficient preference Optimization☆26Oct 10, 2025Updated last year
- ☆17Feb 5, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆14Nov 2, 2025Updated 11 months ago
- Advancing search on top of AI agents☆35Jun 9, 2026Updated 4 months ago
- Clue-RAG: Towards Accurate and Cost-Efficient Graph-based RAG via Multi-Partite Graph and Query-Driven Iterative Retrieval☆26Mar 3, 2026Updated 7 months ago
- Look Back to Reason Forward: Revisitable Memory for Long-Context LLM Agents☆45Apr 13, 2026Updated 5 months ago
- ☆17Jun 15, 2026Updated 3 months ago
- The code implementation for TTCS: Test-Time Curriculum Synthesis for Self-Evolving.☆53Apr 22, 2026Updated 5 months ago
- ☆26Aug 23, 2024Updated 2 years ago
- Some microbenchmarks and design docs before commencement☆11Feb 1, 2021Updated 5 years ago
- something for paper agent☆11Dec 18, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆11Aug 20, 2025Updated last year
- (ACL 2025) Divide-Then-Aggregate: An Efficient Tool Learning Method via Parallel Tool Invocation☆12May 21, 2025Updated last year
- 基于InternLm chat 7B大模型基座,构建一个Agent ,可以调用 MMYOLO 工具来完成图像内视觉任务☆11Oct 30, 2024Updated last year
- ☆43May 9, 2025Updated last year
- [ICML 2025] "From Passive to Active Reasoning: Can Large Language Models Ask the Right Questions under Incomplete Information?"☆48Oct 8, 2025Updated last year
- [COLING 2024] SentiCSE: A Sentiment-aware Contrastive Sentence Embedding Framework with Sentiment-guided Textual Similarity☆13May 8, 2024Updated 2 years ago
- An End-to-End Benchmarking Framework for Retrieval-Augmented Generation Systems☆31Mar 13, 2026Updated 6 months ago
- Run GreenBitAI's Quantized LLMs on Apple Devices with MLX☆31Oct 3, 2026Updated last week
- ☆10Nov 1, 2021Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Metal Max fan game made in Unity☆17Jun 5, 2021Updated 5 years ago
- Implementation of Hindsight Differentiable Policy Optimization, as described in the paper Deep Reinforcement Learning for Inventory Netwo…☆27Nov 19, 2025Updated 10 months ago
- [ICLR 2025] This is the code repo for our ICLR’25 paper "RAG-DDR: Optimizing Retrieval-Augmented Generation Using Differentiable Data Rew…☆56Feb 10, 2025Updated last year
- RL Environment and Benchmark pipeline. Code accompanying paper "Neurophysiologically Realistic Environment for Comparing Adaptive Deep Br…