π Accelerate inference and training of π€ Transformers, Diffusers, TIMM and Sentence Transformers with easy to use hardware optimization tools
β3,473Aug 24, 2026Updated this week
Alternatives and similar repositories for optimum
Users that are interested in optimum are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (iβ¦β9,841Updated this week
- π€ Evaluate: A library for easily evaluating machine learning models and datasets.β2,479Jul 6, 2026Updated last month
- Large Language Model Text Generation Inferenceβ10,889Mar 21, 2026Updated 5 months ago
- Accessible large language models via k-bit quantization for PyTorch.β8,446Updated this week
- π€ PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.β21,605Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Efficient, scalable and enterprise-grade CPU/GPU inference server for π€ Hugging Face transformer models πβ1,689Oct 23, 2024Updated last year
- Transformer related optimization, including BERT, GPTβ6,447Mar 27, 2024Updated 2 years ago
- Simple, safe way to store and distribute tensorsβ3,879Updated this week
- π€ The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation toolsβ21,876Updated this week
- Fast and memory-efficient exact attentionβ24,801Updated this week
- π€ Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.β34,402Updated this week
- π₯ Fast State-of-the-Art Tokenizers optimized for Research and Productionβ11,000Updated this week
- Train transformer language models with reinforcement learning.β19,175Updated this week
- Hackable and optimized Transformers building blocks, supporting a composable construction.β10,544Aug 7, 2026Updated 3 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- MII makes low-latency and high-throughput inference possible, powered by DeepSpeed.β2,111Jun 30, 2025Updated last year
- Efficient few-shot learning with Sentence Transformersβ2,785May 26, 2026Updated 3 months ago
- A pytorch quantization backend for optimumβ1,054Updated this week
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.β43,019Updated this week
- TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizatβ¦β14,497Updated this week
- π€ Optimum Intel: Accelerate inference with Intel optimization toolsβ616Updated this week
- The Triton Inference Server provides an optimized cloud and edge inferencing solution.β10,949Updated this week
- An easy-to-use LLMs quantization package with user-friendly apis, based on GPTQ algorithm.β5,070Apr 11, 2025Updated last year
- ONNX Runtime: cross-platform, high performance ML inferencing and training acceleratorβ21,674Updated this week
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, β¦β2,706Updated this week
- A blazing fast inference solution for text embeddings modelsβ5,028Jul 24, 2026Updated last month
- Development repository for the Triton language and compilerβ20,037Updated this week
- PyTorch extensions for high performance and large scale training.β3,407Apr 26, 2025Updated last year
- Supercharge Your Model Trainingβ5,496Apr 29, 2026Updated 4 months ago
- State-of-the-Art Embeddings, Retrieval, and Rerankingβ19,048Updated this week
- PyTorch native quantization for training and inferenceβ2,959Updated this week
- A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hβ¦β3,509Updated this week
- Foundation Architecture for (M)LLMsβ3,138Apr 11, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Kernl lets you run PyTorch transformer models several times faster on GPU with a single line of code, and is designed to be easily hackabβ¦β1,584Jan 28, 2026Updated 7 months ago
- A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Autoβ¦β18,355Updated this week
- ποΈ A unified multi-backend utility for benchmarking Transformers, Timm, PEFT, Diffusers and Sentence-Transformers with full support of Oβ¦β340May 26, 2026Updated 3 months ago
- Freeing data processing from scripting madness by providing a set of platform-agnostic customizable pipeline processing blocks.β3,310Aug 13, 2026Updated 2 weeks ago
- Minimalistic large language model 3D-parallelism trainingβ2,804May 26, 2026Updated 3 months ago
- Fast inference engine for Transformer modelsβ4,651Aug 16, 2026Updated last week
- Ongoing research training transformer models at scaleβ17,664Updated this week