[NeurIPS'25] Router-R1: Teaching LLMs Multi-Round Routing and Aggregation via Reinforcement Learning
☆150Dec 30, 2025Updated 8 months ago
Alternatives and similar repositories for Router-R1
Users that are interested in Router-R1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2025] "GraphRouter: A Graph-based Router for LLM Selections", Tao Feng, Yanzhen Shen, Jiaxuan You☆79Dec 30, 2025Updated 8 months ago
- [Findings@ACL'26] LLMRouterBench: A Massive Benchmark and Unified Framework for LLM Routing☆99Apr 6, 2026Updated 4 months ago
- [ICML'26] Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory☆23Jun 10, 2026Updated 2 months ago
- [TMLR 2026]: "FusionFactory: Fusing LLM Capabilities with Routing Data", Tao Feng, Haozhen Zhang, Zijie Lei, Pengrui Han, Mostofa Patwary…☆22Dec 30, 2025Updated 8 months ago
- Repo for EmbedLLM: Learning Compact Representations of Large Language Models☆34Sep 25, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ACL'25 Main] Graph of Records: Boosting Retrieval Augmented Generation for Long-context Summarization with Graphs☆41May 26, 2025Updated last year
- LLMRouter: An Open-Source Library for LLM Routing☆2,636Aug 20, 2026Updated last week
- ☆141Oct 29, 2025Updated 10 months ago
- [ICLR 2025] "GraphEval: A Lightweight Graph-Based LLM Framework for Idea Evaluation", Tao Feng, Yihang Sun, Jiaxuan You☆18Mar 18, 2025Updated last year
- ☆17Jun 10, 2025Updated last year
- [ICML 2025]"Graph World Model", Tao Feng, Yexin Wu, Guanyu Lin, Jiaxuan You☆47Sep 20, 2025Updated 11 months ago
- ☆15May 27, 2025Updated last year
- [ACL 2026] R-Search: Empowering LLM Reasoning with Search via Multi-Reward Reinforcement Learning☆36Aug 11, 2026Updated 2 weeks ago
- Code and Data for "FaithfulRAG: Fact-Level Conflict Modeling for Context-Faithful Retrieval-Augmented Generation" (ACL25)☆39Oct 26, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆16Feb 26, 2026Updated 6 months ago
- A curated list of Awesome-LLM-Ensemble papers for the survey "Harnessing Multiple Large Language Models: A Survey on LLM Ensemble"☆254Updated this week
- [EMNLP 2025] DiagramEval: Evaluating LLM-Generated Diagrams via Graphs☆17Nov 1, 2025Updated 9 months ago
- [DAI 2025] Beyond GPT-5: Making LLMs Cheaper and Better via Performance–Efficiency Optimized Routing☆221Dec 11, 2025Updated 8 months ago
- xRouter: Training Cost-Aware LLMs Orchestration System via Reinforcement Learning☆33Jun 2, 2026Updated 2 months ago
- ☆73Oct 23, 2025Updated 10 months ago
- The offical repo for "Parallel-R1: Towards Parallel Thinking via Reinforcement Learning"☆261Feb 4, 2026Updated 6 months ago
- Graph-based experience memory for LLM reward prediction with limited labels. 20% labels → 97.3% Oracle.☆19Mar 24, 2026Updated 5 months ago
- ☆156Jan 21, 2026Updated 7 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [EMNLP26] Welcome! 😊 This is the official code release of EviNote-RAG, and we’re happy to share it with the community.☆49Updated this week
- Opensource code for ICML 2026 poster☆16Nov 26, 2025Updated 9 months ago
- ☆27Feb 28, 2025Updated last year
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,343Nov 13, 2025Updated 9 months ago
- RouterArena: An open framework for evaluating LLM routers with standardized datasets, metrics, an automated framework, and a live leaderb…☆126Aug 20, 2026Updated last week
- PoC for "SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning" [NeurIPS '25]☆75Oct 2, 2025Updated 10 months ago
- Code for 'Answer Matching Outperforms Multiple Choice for Language Model Evaluation' paper☆19Jul 4, 2025Updated last year
- [NeurIPS 2024 Oral] "Bayesian-Guided Label Mapping for Visual Reprogramming"☆12Dec 20, 2024Updated last year
- [ICML'25 Oral] Multi-agent Architecture Search via Agentic Supernet☆282Nov 13, 2025Updated 9 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆15Nov 18, 2025Updated 9 months ago
- 🧠Plan-and-Budget: Training-free test-time reasoning framework for adaptive token allocation in large language models (ICLR 2026).☆15Mar 2, 2026Updated 5 months ago
- [TMLR 2026] A Searching-based Agent Model for Open-Domain Open-Ended Question Answering☆39Jun 20, 2025Updated last year
- COLM2026☆38Jul 9, 2026Updated last month
- On Policy Distillation Build on top of Verl☆97May 25, 2026Updated 3 months ago
- [AAAI 2025] Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks☆12Jun 19, 2025Updated last year
- The implementation for CIKM 2024: Towards Completeness-Oriented Tool Retrieval for Large Language Models.☆26Nov 6, 2024Updated last year