llm-d benchmark scripts and tooling
☆62Jul 28, 2026Updated this week
Alternatives and similar repositories for llm-d-benchmark
Users that are interested in llm-d-benchmark are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Simplified model deployment on llm-d☆29Jul 2, 2025Updated last year
- Incubating P/D sidecar for llm-d☆17Nov 13, 2025Updated 8 months ago
- GenAI inference performance benchmarking tool☆214Updated this week
- Distributed KV cache scheduling & offloading libraries☆165Updated this week
- llm-d Router: The intelligent entry point for inference requests☆272Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- helm charts for deploying models with llm-d☆31Updated this week
- A lightweight, configurable, and real-time simulator designed to mimic the behavior of vLLM without the need for GPUs or running actual h…☆172Updated this week
- llm-d helm charts and deployment examples☆59May 1, 2026Updated 2 months ago
- Kubernetes controllers for fast model actuation using vLLM sleep/wake and launcher-based model swapping☆16Updated this week
- ☆25Updated this week
- Variant optimization autoscaler for distributed inference workloads☆52Updated this week
- ☆22Mar 11, 2026Updated 4 months ago
- Auto-tuning for vllm. Getting the best performance out of your LLM deployment (vllm+guidellm+optuna)☆64Jun 12, 2026Updated last month
- Gateway API Inference Extension☆725Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Inference payload processor for llm-d☆16Updated this week
- ☆110Jul 21, 2025Updated last year
- Let my Claude talk to yours.☆30Jul 22, 2026Updated last week
- Inference Platform Simulation☆21Updated this week
- label ALL kubectl, kustomize, and helm objects, inline, without extra steps.(including namespaces and CRDs)☆15Apr 22, 2024Updated 2 years ago
- Stateful API logic for agentic applications using vLLM☆56Updated this week
- Performance dashboards from the Perf & Scale team☆20Jul 17, 2026Updated last week
- A tool to detect infrastructure issues on cloud native AI systems☆54Sep 18, 2025Updated 10 months ago
- Cloud Native Benchmarking of Foundation Models☆46Jul 31, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Offline optimization of your disaggregated Dynamo graph☆377Updated this week
- Discover ingress-nginx usage and auto-generate Gateway API migration plans before ingress-nginx reaches end-of-life (March 2026).☆16Nov 26, 2025Updated 8 months ago
- Community maintained hardware plugin for vLLM on Spyre☆52Updated this week
- An Envoy inspired, ultimate LLM-first gateway for LLM serving and downstream application developers and enterprises☆27Apr 24, 2025Updated last year
- ☆15May 28, 2024Updated 2 years ago
- kubeadm with core components instrumented to export OpenTelemetry traces (etcd, kube-apiserver, crio)☆13Oct 18, 2022Updated 3 years ago
- A workload for deploying LLM inference services on Kubernetes☆267Updated this week
- An ansible role which configures metrics collection.☆17Updated this week
- General agent evaluation framework☆64Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- d.run website☆18Jul 3, 2026Updated 3 weeks ago
- The main purpose of runtime copilot is to assist with node runtime management tasks such as configuring registries, upgrading versions, i…☆13May 16, 2023Updated 3 years ago
- A framework for designing, executing and analysing experiment campaigns☆56Updated this week
- Automatically scales Kubernetes controllers to zero☆16May 30, 2019Updated 7 years ago
- Solution Service Architecture☆26Jun 5, 2024Updated 2 years ago
- WG Serving☆38Mar 24, 2026Updated 4 months ago
- The Intelligent Inference Scheduler for Large-scale Inference Services.☆68Feb 12, 2026Updated 5 months ago