[NeurIPS'25] Router-R1: Teaching LLMs Multi-Round Routing and Aggregation via Reinforcement Learning
☆149Dec 30, 2025Updated 7 months ago
Alternatives and similar repositories for Router-R1
Users that are interested in Router-R1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2025] "GraphRouter: A Graph-based Router for LLM Selections", Tao Feng, Yanzhen Shen, Jiaxuan You☆75Dec 30, 2025Updated 7 months ago
- [Findings@ACL'26] LLMRouterBench: A Massive Benchmark and Unified Framework for LLM Routing☆91Apr 6, 2026Updated 4 months ago
- [ICML'26] Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory☆22Jun 10, 2026Updated last month
- [TMLR 2026]: "FusionFactory: Fusing LLM Capabilities with Routing Data", Tao Feng, Haozhen Zhang, Zijie Lei, Pengrui Han, Mostofa Patwary…☆22Dec 30, 2025Updated 7 months ago
- [ACL'25 Main] Graph of Records: Boosting Retrieval Augmented Generation for Long-context Summarization with Graphs☆41May 26, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- LLMRouter: An Open-Source Library for LLM Routing☆2,291Updated this week
- ☆138Oct 29, 2025Updated 9 months ago
- [ICLR 2025] "GraphEval: A Lightweight Graph-Based LLM Framework for Idea Evaluation", Tao Feng, Yihang Sun, Jiaxuan You☆17Mar 18, 2025Updated last year
- ☆17Jun 10, 2025Updated last year
- [ICML 2025]"Graph World Model", Tao Feng, Yexin Wu, Guanyu Lin, Jiaxuan You☆44Sep 20, 2025Updated 10 months ago
- ☆15May 27, 2025Updated last year
- [ACL 2026] R-Search: Empowering LLM Reasoning with Search via Multi-Reward Reinforcement Learning☆35Jan 4, 2026Updated 7 months ago
- Code and Data for "FaithfulRAG: Fact-Level Conflict Modeling for Context-Faithful Retrieval-Augmented Generation" (ACL25)☆39Oct 26, 2025Updated 9 months ago
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆16Feb 26, 2026Updated 5 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [NeurIPS 2025] Search and Refine During Think: Facilitating Knowledge Refinement for Improved Retrieval-Augmented Reasoning☆143Jun 25, 2026Updated last month
- A curated list of Awesome-LLM-Ensemble papers for the survey "Harnessing Multiple Large Language Models: A Survey on LLM Ensemble"☆254Updated this week
- [EMNLP 2025] DiagramEval: Evaluating LLM-Generated Diagrams via Graphs☆17Nov 1, 2025Updated 9 months ago
- The code for the paper ROUTERBENCH: A Benchmark for Multi-LLM Routing System☆174Jun 13, 2024Updated 2 years ago
- xRouter: Training Cost-Aware LLMs Orchestration System via Reinforcement Learning☆33Jun 2, 2026Updated 2 months ago
- ☆73Oct 23, 2025Updated 9 months ago
- The offical repo for "Parallel-R1: Towards Parallel Thinking via Reinforcement Learning"☆261Feb 4, 2026Updated 6 months ago
- Graph-based experience memory for LLM reward prediction with limited labels. 20% labels → 97.3% Oracle.☆19Mar 24, 2026Updated 4 months ago
- ☆154Jan 21, 2026Updated 6 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ACL-2026] MMSearch-R1 is an end-to-end RL framework that enables LMMs to perform on-demand, multi-turn search with real-world multimodal…☆477Apr 7, 2026Updated 4 months ago
- Welcome! 😊 This is the official code release of EviNote-RAG, and we’re happy to share it with the community.☆48Jun 4, 2026Updated 2 months ago
- Opensource code for ICML 2026 poster☆15Nov 26, 2025Updated 8 months ago
- ☆26Feb 28, 2025Updated last year
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,278Nov 13, 2025Updated 8 months ago
- PoC for "SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning" [NeurIPS '25]☆75Oct 2, 2025Updated 10 months ago
- Code for 'Answer Matching Outperforms Multiple Choice for Language Model Evaluation' paper☆19Jul 4, 2025Updated last year
- [NeurIPS 2024 Oral] "Bayesian-Guided Label Mapping for Visual Reprogramming"☆12Dec 20, 2024Updated last year
- [ICML'25 Oral] Multi-agent Architecture Search via Agentic Supernet☆280Nov 13, 2025Updated 8 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Nov 18, 2025Updated 8 months ago
- 🧠Plan-and-Budget: Training-free test-time reasoning framework for adaptive token allocation in large language models (ICLR 2026).☆15Mar 2, 2026Updated 5 months ago
- [TMLR 2026] A Searching-based Agent Model for Open-Domain Open-Ended Question Answering☆38Jun 20, 2025Updated last year
- COLM2026☆36Jul 9, 2026Updated last month
- On Policy Distillation Build on top of Verl☆94May 25, 2026Updated 2 months ago
- [AAAI 2025] Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks☆12Jun 19, 2025Updated last year
- The implementation for CIKM 2024: Towards Completeness-Oriented Tool Retrieval for Large Language Models.☆26Nov 6, 2024Updated last year