Incubating P/D sidecar for llm-d
☆17Nov 13, 2025Updated 10 months ago
Alternatives and similar repositories for llm-d-routing-sidecar
Users that are interested in llm-d-routing-sidecar are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Helm charts for llm-d☆52Jul 22, 2025Updated last year
- Distributed KV cache scheduling & offloading libraries☆177Updated this week
- llm-d benchmark scripts and tooling☆67Updated this week
- Simplified model deployment on llm-d☆28Jul 2, 2025Updated last year
- Kubernetes controllers for fast model actuation using vLLM sleep/wake and launcher-based model swapping☆23Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- llm-d Router: The intelligent entry point for inference requests☆334Updated this week
- A lightweight, configurable, and real-time simulator designed to mimic the behavior of vLLM without the need for GPUs or running actual h…☆198Updated this week
- Inference payload processor for llm-d☆19Updated this week
- ☆29Updated this week
- label ALL kubectl, kustomize, and helm objects, inline, without extra steps.(including namespaces and CRDs)☆15Apr 22, 2024Updated 2 years ago
- llm-d helm charts and deployment examples☆59May 1, 2026Updated 4 months ago
- Karmada功能、特性及源码分析☆16Mar 24, 2023Updated 3 years ago
- d.run website☆18Updated this week
- The main purpose of runtime copilot is to assist with node runtime management tasks such as configuring registries, upgrading versions, i…☆13May 16, 2023Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Kubernetes operator for bpfman☆38Jul 14, 2026Updated last month
- Automatically scales Kubernetes controllers to zero☆16May 30, 2019Updated 7 years ago
- Build and deploy Node.js application on Kubernetes☆16Sep 17, 2025Updated 11 months ago
- caniuse.com, but for kubernetes☆27Dec 25, 2024Updated last year
- 中国开发者活动日程(关注点:开源、开发者、云原生)☆30Updated this week
- ☆22Mar 11, 2026Updated 6 months ago
- Variant optimization autoscaler for distributed inference workloads☆57Updated this week
- GenAI inference performance benchmarking tool☆239Updated this week
- Gateway API Inference Extension☆767Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Ingress node firewall implements Kubernetes operator to provision stateless ingress node level firewall rules, stateless ingress node fir…☆77Sep 1, 2026Updated last week
- Operator for managing Node Feature Discovery deployment☆78Mar 12, 2026Updated 6 months ago
- Main rossoctl repo - installer, UI and docs☆300Updated this week
- Let my Claude talk to yours.☆30Aug 5, 2026Updated last month
- CLI for the Serverless Supercomputer☆25Sep 17, 2025Updated 11 months ago
- Operator for the mutating admission webhook for ClusterResourceOverride☆19Updated this week
- Prototypes and experiments for WG Device Management.☆16May 21, 2026Updated 3 months ago
- Kubernetes APIServer 高性能代理组件,代理 APIServer 的 List 请求,其它类型的请求会直接反向代理到原生 APIServer。 CKube 还额外支持了分页、搜索和索引等功能。 并且,CKube 100% 兼容原生 kubectl 和 ku…☆19Sep 16, 2022Updated 3 years ago
- helm repo add daocloud https://daocloud.github.io/dce-charts-repackage/☆13Updated this week
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆15Mar 6, 2025Updated last year
- 🧘 Extensive LLM endpoints, expended capabilities through your favorite protocols, 🕸️ GraphQL, ↔️ gRPC, ♾️ WebSocket. Extended SOTA supp…☆19Updated this week
- ☆21Updated this week
- Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, T…☆508Updated this week
- Mix kubebuilder and code-generator example☆22Jul 3, 2020Updated 6 years ago
- Experimental DRA driver bringing CNI closer to Kubernetes☆45Oct 1, 2025Updated 11 months ago
- NVIDIA Inference Benchmarks provide recipes in ready-to-use templates for evaluating platform speed. Validate your platform across speci…☆88Updated this week