Benchmark and optimize LLM inference across frameworks with ease
☆197Jul 14, 2026Updated last week
Alternatives and similar repositories for llm-optimizer
Users that are interested in llm-optimizer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆12Jul 5, 2023Updated 3 years ago
- Simple dependency injection framework for Python☆21Jul 14, 2026Updated last week
- Genai-bench is a powerful benchmark tool designed for comprehensive token-level performance evaluation of large language model (LLM) serv…☆314Updated this week
- how to build a sentence embedding application using BentoML☆15Jul 14, 2026Updated last week
- 🐳 Build OCI images for Bentos in k8s☆19Jul 14, 2026Updated last week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Evaluate and Enhance Your LLM Deployments for Real-World Inference Needs☆1,429Updated this week
- Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, T…☆482Updated this week
- Auto-tuning for vllm. Getting the best performance out of your LLM deployment (vllm+guidellm+optuna)☆64Jun 12, 2026Updated last month
- ☆151Updated this week
- Checkpoint-engine is a simple middleware to update model weights in LLM inference engines☆982Jul 4, 2026Updated 3 weeks ago
- Gateway API Inference Extension☆723Updated this week
- ArcticInference: vLLM plugin for high-throughput, low-latency inference☆462Jul 14, 2026Updated last week
- Achieve state of the art inference performance with modern accelerators on Kubernetes☆3,875Updated this week
- MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline Parallelism☆18Nov 18, 2025Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- llm-d benchmark scripts and tooling☆62Updated this week
- Intel® AI for Enterprise Inference optimizes AI inference services on Intel hardware using Kubernetes Orchestration. It automates LLM mod…☆44Jul 8, 2026Updated 2 weeks ago
- ☆13Nov 1, 2024Updated last year
- ☆29Jul 14, 2026Updated last week
- A self-evolving personal AI assistant.☆39Mar 13, 2026Updated 4 months ago
- Kubernetes APIServer 高性能代理组件,代理 APIServer 的 List 请求,其它类型的请求会直接反向代理到原生 APIServer。 CKube 还额外支持了分页、搜索和索引等功能。 并且,CKube 100% 兼容原生 kubectl 和 ku…☆19Sep 16, 2022Updated 3 years ago
- The Argo Rollouts plugin implementing the Contour HTTPProxy traffic control in progressive delivery scenarios.☆19Nov 26, 2024Updated last year
- WaveRNN Vocoder + TTS☆11Nov 20, 2021Updated 4 years ago
- Incubating P/D sidecar for llm-d☆17Nov 13, 2025Updated 8 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Helm charts for llm-d☆52Jul 22, 2025Updated last year
- AIPerf is a comprehensive benchmarking tool that measures the performance of generative AI models served by your preferred inference solu…☆468Updated this week
- Applying Machine Learning methodologies in search of novel MOF's and battery materials.☆14May 31, 2023Updated 3 years ago
- ☆26Jul 14, 2026Updated last week
- vLLM’s reference system for K8S-native cluster-wide deployment with community-driven performance optimization☆2,474Updated this week
- ☆330Updated this week