GPU environment and cluster management with LLM support
☆660May 16, 2024Updated 2 years ago
Alternatives and similar repositories for genv
Users that are interested in genv are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A top-like tool for monitoring GPUs in a cluster☆85Feb 14, 2024Updated 2 years ago
- Translation layer that maps any Kubernetes framework's Custom Resource Definitions (CRDs) into a standardized, generic structure.☆60Updated this week
- ☆336Aug 3, 2026Updated last week
- GPU Environment Management for JupyterLab☆26Feb 19, 2024Updated 2 years ago
- KAI Scheduler is an open source Kubernetes Native scheduler for AI workloads at large scale☆1,447Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆304Aug 3, 2026Updated last week
- Tensors, for human consumption☆1,391Apr 9, 2026Updated 4 months ago
- ClearML Fractional GPU - Run multiple containers on the same GPU with driver level memory limitation ✨ and compute time-slicing☆94Mar 12, 2026Updated 4 months ago
- ☆813Updated this week
- The AI Compute Platform for frontier teams. SkyPilot turns fragmented AI compute into one AI supercomputer, so frontier AI teams build cu…☆10,465Updated this week
- Vendor-agnostic orchestration for training, inference and agentic workloads across NVIDIA, AMD, TPU, and Tenstorrent on clouds, Kubernete…☆2,209Updated this week
- DRA Driver for NVIDIA GPUs☆687Updated this week
- GPUd automates monitoring, diagnostics, and issue identification for GPUs☆487Updated this week
- Practical GPU Sharing Without Memory Size Constraints☆315Mar 28, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A comprehensive Helm chart for monitoring GPU resources in Kubernetes clusters. This tool provides real-time visibility into GPU allocati…☆33Jun 30, 2026Updated last month
- NVIDIA GPU Operator creates, configures, and manages GPUs in Kubernetes☆2,828Updated this week
- Kubernetes enhancements for Network Topology Aware Gang Scheduling & Autoscaling☆251Updated this week
- A transparent, in-container GPU resource controller that enforces memory and compute limits by intercepting CUDA calls without applicatio…☆321Aug 3, 2026Updated last week
- A Datacenter Scale Distributed Inference Serving Framework☆7,729Updated this week
- ☆16Jul 20, 2026Updated 3 weeks ago
- GPU plugin to the node feature discovery for Kubernetes☆309May 27, 2024Updated 2 years ago
- Heterogeneous GPU Sharing on Kubernetes☆4,289Updated this week
- LeaderWorkerSet: An API for deploying a group of pods as a unit of replication☆783Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A JupyterLab extension for displaying dashboards of GPU usage.☆680Updated this week
- The NVIDIA Driver Manager is a Kubernetes component which assist in seamless upgrades of NVIDIA Driver on each node of the cluster.☆55Updated this week
- An Envoy inspired, ultimate LLM-first gateway for LLM serving and downstream application developers and enterprises☆27Apr 24, 2025Updated last year
- The Triton Inference Server provides an optimized cloud and edge inferencing solution.☆10,913Updated this week
- ☆12Aug 27, 2024Updated last year
- ☆207Jul 12, 2026Updated 3 weeks ago
- Supercharge Your Model Training☆5,494Apr 29, 2026Updated 3 months ago
- NVIDIA device plugin for Kubernetes☆3,844Updated this week
- Tools for building GPU clusters☆1,466Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Kubernetes Operator for MPI-based applications (distributed training, HPC, etc.)☆531Updated this week
- elastic-gpu-scheduler is a Kubernetes scheduler extender for GPU resources scheduling.☆147Nov 21, 2022Updated 3 years ago
- Python client for the Run:ai REST API☆26Dec 15, 2025Updated 7 months ago
- Simple, safe way to store and distribute tensors☆3,851Updated this week
- AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-te…☆1,240Jul 31, 2026Updated last week
- GPU Sharing Scheduler for Kubernetes Cluster☆1,531Dec 29, 2023Updated 2 years ago
- Containers for machine learning☆9,454Updated this week