RouterArena: An open framework for evaluating LLM routers with standardized datasets, metrics, an automated framework, and a live leaderboard.
☆119Aug 5, 2026Updated 2 weeks ago
Alternatives and similar repositories for RouterArena
Users that are interested in RouterArena are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17Apr 22, 2026Updated 3 months ago
- This is the repo for remote direct memory introspection.☆25Jun 21, 2023Updated 3 years ago
- [ICML 2026] Decoding Tree Sketching (DTS): a training-free & model agonistic & plug-in framework for LLM parallel reasoning.☆72May 12, 2026Updated 3 months ago
- Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond☆1,136Updated this week
- ☆22Oct 3, 2024Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Zodiac: Unearthing Semantic Checks for Cloud Infrastructure-as-Code Programs, SOSP 2024☆15Nov 28, 2024Updated last year
- [ICLR 2025] "GraphRouter: A Graph-based Router for LLM Selections", Tao Feng, Yanzhen Shen, Jiaxuan You☆75Dec 30, 2025Updated 7 months ago
- 山东大学软件学院历年考试真题及参考答案☆14Jun 27, 2022Updated 4 years ago
- [Findings@ACL'26] LLMRouterBench: A Massive Benchmark and Unified Framework for LLM Routing☆95Apr 6, 2026Updated 4 months ago
- Repo for EmbedLLM: Learning Compact Representations of Large Language Models☆34Sep 25, 2025Updated 10 months ago
- An enhanced version of BMv2 switches that support changing the P4 program at runtime with no service interruption.☆19Apr 12, 2023Updated 3 years ago
- Mitigating Network Covert Channels while Preserving Performance☆19Aug 12, 2020Updated 6 years ago
- [ACM APNet 2025 Best Paper Award] A high-performance and easy-to-use RDMA library, called SnowRDMA.☆26Jul 15, 2024Updated 2 years ago
- [NeurIPS'25] Router-R1: Teaching LLMs Multi-Round Routing and Aggregation via Reinforcement Learning☆149Dec 30, 2025Updated 7 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- SwiftRDMA -- Exposing RDMA NIC Resources for Software-Defined RDMA Scheduling☆53Jun 9, 2026Updated 2 months ago
- Accelerating Large-Scale Reasoning Model Inference with Sparse Self-Speculative Decoding☆119Dec 2, 2025Updated 8 months ago
- Ripple: A Programmable, Decentralized Link-Flooding Defense Against Adaptive Adversaries☆25May 21, 2022Updated 4 years ago
- An Automated Performance Optimization Framework for P4-Programmable SmartNICs☆28Nov 18, 2023Updated 2 years ago
- Poise source code repo☆12Aug 12, 2020Updated 6 years ago
- ☆24Jan 18, 2026Updated 7 months ago
- [ICML 2024] Watermarks in the Sand: Impossibility of Strong Watermarking for Generative Models☆27Sep 12, 2024Updated last year
- ☆27Dec 20, 2021Updated 4 years ago
- A WebUI for Side-by-Side Comparison of Media (Images/Videos) Across Multiple Folders☆26Feb 21, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Sardeenz is a proof-of-concept application that allows you to load more than one model on a given GPU. It allows you to add more and more…☆61Updated this week
- A Multi-Policy, Multi-Agent RL Training Framework☆32Jun 16, 2026Updated 2 months ago
- Following the same workflows as Kubernetes. Widely used in InftyAI community.☆12May 31, 2026Updated 2 months ago
- Provides deploy scripts and CSI for Lustre.☆14Apr 13, 2026Updated 4 months ago
- This guide is beginner-friendly, project-driven, and laser-focused on the commands & concepts you will actually use while working with Do…☆16Dec 20, 2025Updated 7 months ago
- KV Cache Compression, But What Must We Give in Return? A Comprehensive Benchmark of Long Context Capable Approaches. EMNLP Findings 2024☆90Feb 27, 2025Updated last year
- ☆24Dec 6, 2025Updated 8 months ago
- Route LLM requests to the best model for the task at hand.☆341Jul 24, 2026Updated 3 weeks ago
- A programmable Mixture-of-Models router for heterogeneous LLM inference☆5,177Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆48Jul 27, 2026Updated 3 weeks ago
- Official code release for Delta Activations: A Representation for Finetuned Large Language Models☆21Sep 5, 2025Updated 11 months ago
- A Maximal Mutual Information Criterion for Manipulation Concept Discovery☆14Sep 26, 2024Updated last year
- ☆17Jun 26, 2025Updated last year
- Research prototype of PRISM — a cost-efficient multi-LLM serving system with flexible time- and space-based GPU sharing.