Module to Automatically maximize the utilization of GPU resources in a Kubernetes cluster through real-time dynamic partitioning and elastic quotas - Effortless optimization at its finest!
☆686Apr 21, 2024Updated 2 years ago
Alternatives and similar repositories for nos
Users that are interested in nos are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- NVIDIA device plugin for Kubernetes☆49Feb 16, 2024Updated 2 years ago
- NVIDIA device plugin for Kubernetes☆3,847Updated this week
- DRA Driver for NVIDIA GPUs☆688Updated this week
- GPU Sharing Scheduler for Kubernetes Cluster☆1,531Dec 29, 2023Updated 2 years ago
- Kubernetes-native Job Queueing☆2,793Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- KAI Scheduler is an open source Kubernetes Native scheduler for AI workloads at large scale☆1,447Updated this week
- A collection of libraries to optimise AI model performances☆8,329Jul 22, 2024Updated 2 years ago
- NVIDIA GPU Operator creates, configures, and manages GPUs in Kubernetes☆2,828Updated this week
- Heterogeneous GPU Sharing on Kubernetes☆4,304Updated this week
- A Cloud Native Batch System (Project under CNCF)☆5,845Updated this week
- AWS virtual gpu device plugin provides capability to use smaller virtual gpus for your machine learning inference workloads☆203Nov 22, 2023Updated 2 years ago
- A kubernetes operator for creating and managing a cache of container images directly on the cluster worker nodes, so application pods sta…☆1,371Jul 17, 2026Updated 3 weeks ago
- Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes☆5,783Updated this week
- Kubectl Sockperf plugin - Latency Measurement in Kubernetes☆22Nov 26, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- MIG Partition Editor for NVIDIA GPUs☆260Updated this week
- JobSet: a k8s native API for distributed ML training and HPC workloads☆335Updated this week
- AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-te…☆1,240Jul 31, 2026Updated last week
- LeaderWorkerSet: An API for deploying a group of pods as a unit of replication☆786Updated this week
- A Kubernetes plugin that gives context to what is restarting in your Kubernetes cluster☆155Sep 10, 2025Updated 11 months ago
- Repository for out-of-tree scheduler plugins based on scheduler framework.☆1,315Jul 9, 2026Updated last month
- Kpad is a simple multiplatform terminal editor born to edit kubernetes declarative manifest yaml files.☆44Oct 10, 2023Updated 2 years ago
- Run Slurm in Kubernetes☆412Updated this week
- Practical GPU Sharing Without Memory Size Constraints☆316Mar 28, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Automatically taint nodes and evict pods based on cpu pressure☆51Dec 23, 2022Updated 3 years ago
- Multi-tenancy and policy-based framework for Kubernetes.☆2,151Updated this week
- Multi-cluster Kubernetes usage analytics for CPU, Memory, and GPU — track costs and optimize cluster resources☆63Mar 18, 2025Updated last year
- Resource-adaptive cluster scheduler for deep learning training.☆459Mar 5, 2023Updated 3 years ago
- Kubernetes Operator for MPI-based applications (distributed training, HPC, etc.)☆531Updated this week
- A Kubernetes controller for automatically optimizing pod requests based on their continuous usage. VPA alternative that can work with HPA…☆207Feb 9, 2024Updated 2 years ago
- ☆905Apr 2, 2024Updated 2 years ago
- Enable dynamic and seamless Kubernetes multi-cluster topologies☆1,474Updated this week
- knavigator is a development, testing, and optimization toolkit for AI/ML scheduling systems at scale on Kubernetes.☆80Jul 6, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- K8s device plugin for GPU sharing☆100May 10, 2023Updated 3 years ago
- Kubernetes AI Toolchain Operator☆997Updated this week
- A light library to allow changing pod log level without restarting the pod☆12Jul 29, 2023Updated 3 years ago
- ☆14Jan 11, 2023Updated 3 years ago
- A Topology-Aware Custom Scheduler For Kubernetes☆65Jul 5, 2023Updated 3 years ago
- The AI Compute Platform for frontier teams. SkyPilot turns fragmented AI compute into one AI supercomputer, so frontier AI teams build cu…☆10,469Updated this week
- Holistic job manager on Kubernetes☆117Feb 20, 2024Updated 2 years ago