⚡VoltaML is a lightweight library to convert and run your ML/DL deep learning models in high performance inference runtimes like TensorRT, TorchScript, ONNX and TVM.
☆1,176Nov 30, 2022Updated 3 years ago
Alternatives and similar repositories for voltaML
Users that are interested in voltaML are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Beautiful and Easy to use Stable Diffusion WebUI☆997Jun 19, 2024Updated 2 years ago
- Kernl lets you run PyTorch transformer models several times faster on GPU with a single line of code, and is designed to be easily hackab…☆1,584Jan 28, 2026Updated 7 months ago
- AITemplate is a Python framework which renders neural network into high performance CUDA/HIP C++ code. Specialized for FP16 TensorCore (N…☆4,724Aug 7, 2026Updated 3 weeks ago
- Efficient, scalable and enterprise-grade CPU/GPU inference server for 🤗 Hugging Face transformer models 🚀☆1,690Oct 23, 2024Updated last year
- 🐶 A tool to package, serve, and deploy any ML model on any platform. Archived to be resurrected one day🤞☆718Sep 13, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Real-time inference for Stable Diffusion - 0.88s latency. Covers AITemplate, nvFuser, TensorRT, FlashAttention. Join our Discord communty…☆556Dec 4, 2023Updated 2 years ago
- PyQt6 GUI to queue and render images and videos using ComfyUI Workflows☆22Jul 20, 2026Updated last month
- Supercharge Your Model Training☆5,493Apr 29, 2026Updated 4 months ago
- 🚀 Accelerate inference and training of 🤗 Transformers, Diffusers, TIMM and Sentence Transformers with easy to use hardware optimization…☆3,479Aug 24, 2026Updated last week
- Fast finetuning using a booster model that puts the initial state to a local minimum☆113Aug 29, 2023Updated 3 years ago
- Efficient few-shot learning with Sentence Transformers☆2,788Updated this week
- PyTriton is a Flask/FastAPI-like interface that simplifies Triton's deployment in Python environments.☆848Aug 13, 2025Updated last year
- Official Implementation of Paella https://arxiv.org/abs/2211.07292v2☆748Oct 4, 2023Updated 2 years ago
- Large-scale model inference.☆628Sep 12, 2023Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Containers for machine learning☆9,468Updated this week
- A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)☆4,754Jan 8, 2024Updated 2 years ago
- The WeightWatcher tool for predicting the accuracy of Deep Neural Networks☆1,774May 11, 2026Updated 3 months ago
- 🛠️ Tools for Transformers compression using PyTorch Lightning ⚡☆85Aug 4, 2026Updated last month
- Sparsity-aware deep learning inference runtime for CPUs☆3,159Jun 2, 2025Updated last year
- Desktop AI Generator☆101Apr 14, 2023Updated 3 years ago
- The simplest way to serve AI/ML models in production☆1,199Updated this week
- MegCC是一个运行时超轻量,高效,移植简单的深度学习模型编译器☆483Oct 23, 2024Updated last year
- 🦘 Explore multimedia datasets at scale☆1,077Dec 7, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Foundation Architecture for (M)LLMs☆3,139Apr 11, 2024Updated 2 years ago
- https://wavespeed.ai/ Best inference performance optimization framework for HuggingFace Diffusers on NVIDIA GPUs.☆1,304Mar 27, 2025Updated last year
- An open-source efficient deep learning framework/compiler, written in python.☆743Sep 4, 2025Updated last year
- SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, …☆2,707Updated this week
- Python-based research interface for blackbox and hyperparameter optimization, based on the internal Google Vizier Service.☆1,672Updated this week
- General technology for enabling AI capabilities w/ LLMs and MLLMs☆4,467Jul 25, 2026Updated last month
- An open-source ML pipeline development platform☆1,002Jan 9, 2025Updated last year
- An implementation of a server for the Stability AI Stable Diffusion API☆172Jan 6, 2023Updated 3 years ago
- Argilla is a collaboration tool for AI engineers and domain experts to build high-quality datasets☆5,096Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- General fine tuning for Stable Diffusion☆505Apr 30, 2023Updated 3 years ago
- Active Learning for Text Classification in Python☆646May 24, 2026Updated 3 months ago
- An open-source AutoML Library based on PyTorch☆306Updated this week
- Accessible large language models via k-bit quantization for PyTorch.☆8,455Updated this week
- Command Line Interface for Hugging Face Inference Endpoints☆65Apr 10, 2024Updated 2 years ago
- A collection of libraries to optimise AI model performances☆8,328Jul 22, 2024Updated 2 years ago
- ☆33Jan 12, 2024Updated 2 years ago