π Accelerate inference and training of π€ Transformers, Diffusers, TIMM and Sentence Transformers with easy to use hardware optimization tools
β3,498Oct 6, 2026Updated this week
Alternatives and similar repositories for optimum
Users that are interested in optimum are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (iβ¦β9,911Updated this week
- π€ Evaluate: A library for easily evaluating machine learning models and datasets.β2,488Sep 23, 2026Updated 2 weeks ago
- Large Language Model Text Generation Inferenceβ10,880Mar 21, 2026Updated 6 months ago
- Accessible large language models via k-bit quantization for PyTorch.β8,514Sep 7, 2026Updated last month
- π€ PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.β21,772Updated this week
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Efficient, scalable and enterprise-grade CPU/GPU inference server for π€ Hugging Face transformer models πβ1,693Oct 23, 2024Updated last year
- Transformer related optimization, including BERT, GPTβ6,454Mar 27, 2024Updated 2 years ago
- Simple, safe way to store and distribute tensorsβ3,913Updated this week
- π€ The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation toolsβ22,044Updated this week
- Fast and memory-efficient exact attentionβ25,107Updated this week
- π€ Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.β34,697Updated this week
- π₯ Fast State-of-the-Art Tokenizers optimized for Research and Productionβ11,164Updated this week
- Train transformer language models with reinforcement learning.β19,476Updated this week
- Hackable and optimized Transformers building blocks, supporting a composable construction.β10,560Sep 23, 2026Updated 2 weeks ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- MII makes low-latency and high-throughput inference possible, powered by DeepSpeed.β2,114Jun 30, 2025Updated last year
- Efficient few-shot learning with Sentence Transformersβ2,836Updated this week
- A pytorch quantization backend for optimumβ1,052Sep 29, 2026Updated last week
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.β43,214Updated this week
- π€ Optimum Intel: Accelerate inference with Intel optimization toolsβ622Updated this week
- TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizatβ¦β14,780Updated this week
- The Triton Inference Server provides an optimized cloud and edge inferencing solution.β11,059Updated this week
- An easy-to-use LLMs quantization package with user-friendly apis, based on GPTQ algorithm.β5,065Apr 11, 2025Updated last year
- SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, β¦β2,713Updated this week
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ONNX Runtime: cross-platform, high performance ML inferencing and training acceleratorβ22,042Updated this week
- A blazing fast inference solution for text embeddings modelsβ5,078Updated this week
- Development repository for the Triton language and compilerβ20,321Updated this week
- PyTorch extensions for high performance and large scale training.β3,404Apr 26, 2025Updated last year
- Supercharge Your Model Trainingβ5,505Apr 29, 2026Updated 5 months ago
- State-of-the-Art Embeddings, Retrieval, and Rerankingβ19,161Updated this week
- PyTorch native quantization for training and inferenceβ2,995Updated this week
- A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hβ¦β3,571Updated this week
- Foundation Architecture for (M)LLMsβ3,141Apr 11, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Kernl lets you run PyTorch transformer models several times faster on GPU with a single line of code, and is designed to be easily hackabβ¦β1,584Jan 28, 2026Updated 8 months ago
- A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Autoβ¦β18,564Updated this week
- ποΈ A unified multi-backend utility for benchmarking Transformers, Timm, PEFT, Diffusers and Sentence-Transformers with full support of Oβ¦β340Sep 29, 2026Updated last week
- Freeing data processing from scripting madness by providing a set of platform-agnostic customizable pipeline processing blocks.β3,377Updated this week
- Minimalistic large language model 3D-parallelism trainingβ2,834Updated this week
- Fast inference engine for Transformer modelsβ4,701Updated this week
- Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalitiesβ22,228Sep 21, 2026Updated 2 weeks ago