The code for the paper ROUTERBENCH: A Benchmark for Multi-LLM Routing System
☆175Jun 13, 2024Updated 2 years ago
Alternatives and similar repositories for routerbench
Users that are interested in routerbench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Open sourced backend for Martian's LLM Inference Provider Leaderboard☆21Aug 13, 2024Updated 2 years ago
- Tutorial for building LLM router☆256Jul 19, 2024Updated 2 years ago
- ☆13Updated this week
- Framework for Cost-Effective Language Model Choice☆16Dec 12, 2023Updated 2 years ago
- A framework for serving and evaluating LLM routers - save LLM costs without compromising quality☆5,361Aug 10, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [NeurIPS'25] Router-R1: Teaching LLMs Multi-Round Routing and Aggregation via Reinforcement Learning☆149Dec 30, 2025Updated 7 months ago
- FrugalGPT: better quality and lower cost for LLM applications☆282Feb 10, 2025Updated last year
- Radiantloom Email Assist 7B is an email-assistant large language model fine-tuned from Zephyr-7B-Beta, over a custom-curated dataset of 1…☆14Jan 19, 2024Updated 2 years ago
- Official code for the paper: "Metadata Archaeology"☆19May 10, 2023Updated 3 years ago
- [ACL'25] Code for ACL'25 paper "IRT-Router: Effective and Interpretable Multi-LLM Routing via Item Response Theory"☆34Feb 19, 2025Updated last year
- This is the oficial repository for "Safer-Instruct: Aligning Language Models with Automated Preference Data"☆17Feb 22, 2024Updated 2 years ago
- [ICLR 2025] Official Pytorch Implementation of "Mix-LN: Unleashing the Power of Deeper Layers by Combining Pre-LN and Post-LN" by Pengxia…☆30Jul 24, 2025Updated last year
- ☆56Jun 23, 2026Updated last month
- A Python wrapper around HuggingFace's TGI (text-generation-inference) and TEI (text-embedding-inference) servers.☆32Sep 19, 2025Updated 10 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆56Jun 26, 2025Updated last year
- Archon provides a modular framework for combining different inference-time techniques and LMs with just a JSON config file.☆210Mar 7, 2025Updated last year
- ☆105Dec 6, 2024Updated last year
- BenchBench is a Python package to evaluate multi-task benchmarks.☆23Oct 12, 2025Updated 10 months ago
- Reference implementation for Reward-Augmented Decoding: Efficient Controlled Text Generation With a Unidirectional Reward Model☆45Oct 1, 2025Updated 10 months ago
- Tools for merging pretrained large language models.☆19Jun 12, 2024Updated 2 years ago
- Code for Fooling Contrastive Language-Image Pre-trainined Models with CLIPMasterPrints☆15Jan 25, 2026Updated 6 months ago
- Code for the paper "Learning Step-Size Adaptation in CMA-ES"☆13Mar 24, 2023Updated 3 years ago
- A compute framework for building Search, RAG, Recommendations and Analytics over complex (structured+unstructured) data, with ultra-modal…☆12Sep 16, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Extend existing LLMs way beyond the original training length with constant memory usage, without retraining☆735Apr 10, 2024Updated 2 years ago
- PyTorch implementation of paper "Response Length Perception and Sequence Scheduling: An LLM-Empowered LLM Inference Pipeline".☆93May 23, 2023Updated 3 years ago
- Official pytorch implementation of "Towards Practical Plug-and-Play Diffusion Models" in CVPR2023☆22Jul 22, 2023Updated 3 years ago
- PB-LLM: Partially Binarized Large Language Models☆158Nov 20, 2023Updated 2 years ago
- ☆23Jul 10, 2023Updated 3 years ago
- A lightweight library for generating synthetic instruction tuning datasets for your data without GPT.☆828Jul 15, 2025Updated last year
- 언어모델을 학습하기 위한 공개 한국어 instruction dataset들을 모아두었습니다.☆19Jul 16, 2023Updated 3 years ago
- State of What Art? A Call for Multi-Prompt LLM Evaluation☆15Apr 10, 2026Updated 4 months ago
- ☆25Dec 18, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Superfast AI decision making and intelligent processing of multi-modal data.☆3,821Aug 10, 2026Updated last week
- Inference code for Mistral and Mixtral hacked up into original Llama implementation☆367Dec 9, 2023Updated 2 years ago
- ☆19Dec 31, 2025Updated 7 months ago
- Automatically Discovering Fast Parallelization Strategies for Distributed Deep Neural Network Training☆1,898Aug 11, 2026Updated last week
- Versatile framework designed to streamline the integration of your models, as well as those sourced from Hugging Face, into complex progr…☆35Aug 21, 2025Updated 11 months ago
- DeepTrace: A lightweight, scalable real-time diagnostic and analysis tool for distributed training tasks.☆18Nov 4, 2025Updated 9 months ago
- ☆19Feb 28, 2022Updated 4 years ago