A collection of YAML files, Helm Charts, Operator code, and guides to act as an example reference implementation for NVIDIA NIM deployment.
☆240May 15, 2026Updated 2 months ago
Alternatives and similar repositories for nim-deploy
Users that are interested in nim-deploy are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Accelerate your Gen AI with NVIDIA NIM and NVIDIA AI Workbench☆259May 1, 2025Updated last year
- An Operator for deployment and maintenance of NVIDIA NIMs and NeMo microservices in a Kubernetes environment.☆159Updated this week
- 🚀 Use NVIDIA NIMs with Haystack pipelines☆32Sep 4, 2024Updated last year
- ☆64Jul 15, 2026Updated last week
- Source Code and Usage Samples for the Resources hosted in the NVIDIA AI Enterprise AzureML Registry☆21Aug 7, 2024Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Run cloud native workloads on NVIDIA GPUs☆239Jul 14, 2026Updated last week
- Infrastructure as code for GPU accelerated managed Kubernetes clusters.☆60Apr 30, 2025Updated last year
- Generative AI reference workflows optimized for accelerated infrastructure and microservice architecture.☆4,126May 29, 2026Updated last month
- Training NVIDIA NeMo Megatron Large Language Model (LLM) using NeMo Framework on Google Kubernetes Engine☆16Apr 28, 2025Updated last year
- Provides end-to-end model development pipelines for LLMs and Multimodal models that can be launched on-prem or cloud-native.☆521Apr 18, 2025Updated last year
- ☆12Mar 16, 2026Updated 4 months ago
- ☆13Dec 20, 2025Updated 7 months ago
- HyDE based RAG using NVIDIA NIM.☆16Mar 20, 2024Updated 2 years ago
- Easy to use python wrapper around Deepstream Python bindings☆14Oct 11, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- MIG Partition Editor for NVIDIA GPUs☆259Updated this week
- Community examples utilizing NVIDIA NeMo Agent Toolkit.☆29Updated this week
- DRA Driver for NVIDIA GPUs☆675Updated this week
- Triton CLI is an open source command line interface that enables users to create, deploy, and profile models served by the Triton Inferen…☆73Updated this week
- Gateway API Inference Extension☆723Updated this week
- Comprehensive, scalable ML inference architecture using Amazon EKS, leveraging Graviton processors for cost-effective CPU-based inference…☆19Mar 12, 2026Updated 4 months ago
- The NVIDIA NeMo Agent toolkit is an open-source library for efficiently connecting and optimizing teams of AI agents.☆2,520Updated this week
- ☆206Jul 15, 2026Updated last week
- NVIDIA GPU Operator creates, configures, and manages GPUs in Kubernetes☆2,801Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Issues related to MLPerf® Inference policies, including rules and suggested changes☆62Updated this week
- Enhance your knowledge in medical research with the help of LLM and RAG.☆35Oct 24, 2024Updated last year
- A gateway for separating tools from agent code☆15Jun 20, 2026Updated last month
- Custom Scheduler to deploy ML models to TRTIS for GPU Sharing☆12Apr 1, 2020Updated 6 years ago
- Plugins for Sonobuoy☆61May 20, 2025Updated last year
- Experimenting text-embeddings-inference server on both CPU and GPU☆18Oct 25, 2023Updated 2 years ago
- An NVIDIA AI Workbench example project for Retrieval Augmented Generation (RAG)☆369Aug 12, 2025Updated 11 months ago
- Create, List, Update, Delete Amazon EKS clusters. Deploy and manage software on EKS. Run distributed model training and inference example…☆66Jul 8, 2026Updated 2 weeks ago
- Kubernetes Operator, Helm Charts, Ansible Playbooks, and utility scripts for large-scale AIStore deployments on Kubernetes.☆132Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A Datacenter Scale Distributed Inference Serving Framework☆7,553Updated this week
- LeaderWorkerSet: An API for deploying a group of pods as a unit of replication☆767Updated this week
- This repository contains the results and code for the MLPerf™ Training v3.0 benchmark.☆12Aug 10, 2023Updated 2 years ago
- Recipes for reproducing training and serving benchmarks for large machine learning models using GPUs on Google Cloud.☆138Updated this week
- ☆14May 29, 2024Updated 2 years ago
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆27Jul 6, 2026Updated 2 weeks ago
- ☆20Mar 18, 2026Updated 4 months ago