The code for the paper ROUTERBENCH: A Benchmark for Multi-LLM Routing System
☆154Jun 13, 2024Updated last year
Alternatives and similar repositories for routerbench
Users that are interested in routerbench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Open sourced backend for Martian's LLM Inference Provider Leaderboard☆21Aug 13, 2024Updated last year
- Repo for EmbedLLM: Learning Compact Representations of Large Language Models☆29Sep 25, 2025Updated 6 months ago
- A curated list of awesome approaches to AI model routing☆192Mar 24, 2025Updated last year
- Tutorial for building LLM router☆246Jul 19, 2024Updated last year
- ☆12Feb 11, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click and start building anything your business needs.
- Agent that routes to different tools - LLM classifier SDK☆45Jun 25, 2024Updated last year
- Framework for Cost-Effective Language Model Choice☆16Dec 12, 2023Updated 2 years ago
- A framework for serving and evaluating LLM routers - save LLM costs without compromising quality☆4,720Aug 10, 2024Updated last year
- [ICLR 2025] Official Pytorch Implementation of "Mix-LN: Unleashing the Power of Deeper Layers by Combining Pre-LN and Post-LN" by Pengxia…☆29Jul 24, 2025Updated 8 months ago
- Radiantloom Email Assist 7B is an email-assistant large language model fine-tuned from Zephyr-7B-Beta, over a custom-curated dataset of 1…☆14Jan 19, 2024Updated 2 years ago
- Official code for the paper: "Metadata Archaeology"☆19May 10, 2023Updated 2 years ago
- A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in Large Language Models☆110Nov 12, 2025Updated 4 months ago
- ☆56Jun 26, 2025Updated 9 months ago
- ☆19Dec 31, 2025Updated 2 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆104Dec 6, 2024Updated last year
- This is the oficial repository for "Safer-Instruct: Aligning Language Models with Automated Preference Data"☆17Feb 22, 2024Updated 2 years ago
- ☆55Jan 15, 2026Updated 2 months ago
- Distributed multi-agent framework for event-driven, graph-based computation. Elixir/Python, NATS event streaming, modular operator/XCS ar…☆14Nov 4, 2025Updated 4 months ago
- ☆23Dec 18, 2024Updated last year
- A Python wrapper around HuggingFace's TGI (text-generation-inference) and TEI (text-embedding-inference) servers.☆32Sep 19, 2025Updated 6 months ago
- Archon provides a modular framework for combining different inference-time techniques and LMs with just a JSON config file.☆190Mar 7, 2025Updated last year
- Reference implementation for Reward-Augmented Decoding: Efficient Controlled Text Generation With a Unidirectional Reward Model☆45Oct 1, 2025Updated 5 months ago
- Superfast AI decision making and intelligent processing of multi-modal data.☆3,381Mar 12, 2026Updated 2 weeks ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code for Fooling Contrastive Language-Image Pre-trainined Models with CLIPMasterPrints☆15Jan 25, 2026Updated 2 months ago
- A compute framework for building Search, RAG, Recommendations and Analytics over complex (structured+unstructured) data, with ultra-modal…☆12Sep 16, 2024Updated last year
- ☆49Mar 20, 2026Updated last week
- Official pytorch implementation of "Towards Practical Plug-and-Play Diffusion Models" in CVPR2023☆22Jul 22, 2023Updated 2 years ago
- S-LoRA: Serving Thousands of Concurrent LoRA Adapters☆1,903Jan 21, 2024Updated 2 years ago
- Code repo for efficient quantized MoE inference with mixture of low-rank compensators☆35Apr 14, 2025Updated 11 months ago
- PyTorch implementation of paper "Response Length Perception and Sequence Scheduling: An LLM-Empowered LLM Inference Pipeline".☆92May 23, 2023Updated 2 years ago
- BenchBench is a Python package to evaluate multi-task benchmarks.☆18Oct 12, 2025Updated 5 months ago
- A lightweight library for generating synthetic instruction tuning datasets for your data without GPT.☆823Jul 15, 2025Updated 8 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Tools for merging pretrained large language models.☆19Jun 12, 2024Updated last year
- ☆23Jul 10, 2023Updated 2 years ago
- State of What Art? A Call for Multi-Prompt LLM Evaluation☆15Jul 10, 2024Updated last year
- Code for "RADCoT: Retrieval-Augmented Distillation to Specialization Models for Generating Chain-of-Thoughts in Query Expansion", LREC-CO…☆11May 25, 2024Updated last year
- Routing on Random Forest (RoRF)☆239Sep 24, 2024Updated last year
- GiGA Genie INSIDE(G-INSIDE) SDK☆11Jul 31, 2024Updated last year
- US Neighborhood data in GeoJSON format from OpenSource Zillow Neighborhood Boundaries Shapefiles☆11Oct 27, 2016Updated 9 years ago