π Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support
β743Jul 20, 2026Updated this week
Alternatives and similar repositories for Automodel
Users that are interested in Automodel are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Training library for Megatron-based models with bidirectional Hugging Face conversion capabilityβ817Updated this week
- Scalable toolkit for efficient model reinforcementβ1,835Updated this week
- A library for exporting models including NeMo and Hugging Face to optimized inference backends, and deploying them for efficient queryingβ41Jul 7, 2026Updated 2 weeks ago
- Scalable data pre processing and curation toolkit for LLMsβ1,672Updated this week
- VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zooβ2,097Updated this week
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.β1,759Updated this week
- Accelerating MoE with IO and Tile-aware Optimizationsβ732Jul 4, 2026Updated 2 weeks ago
- A tool to configure, launch and manage your machine learning experiments.β251Updated this week
- β209Updated this week
- Best practices for training DeepSeek, Mixtral, Qwen and other MoE models using Megatron Core.β201May 29, 2026Updated last month
- Bridge Megatron-Core to Hugging Face/Reinforcement Learningβ226Jun 15, 2026Updated last month
- Megatron's multi-modal data loaderβ374Updated this week
- β537Updated this week
- slime is an LLM post-training framework for RL Scaling.β7,551Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A Distributed Attention Towards Linear Scalability for Ultra-Long Context, Heterogeneous Data Trainingβ883Updated this week
- Evaluate and improve models and agents using environmentsβ1,059Updated this week
- Train speculative decoding models effortlessly and port them smoothly to SGLang serving.β997Updated this week
- π Efficient implementations for emerging model architecturesβ5,379Updated this week
- A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hβ¦β3,435Updated this week
- β37Aug 7, 2025Updated 11 months ago
- A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculativeβ¦β3,266Updated this week
- Checkpoint-engine is a simple middleware to update model weights in LLM inference enginesβ970Jul 4, 2026Updated 2 weeks ago
- A PyTorch native platform for training generative AI modelsβ5,545Updated this week
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Distributed Compiler based on Triton for Parallel Systemsβ1,494Updated this week
- CUDA kernels for linear attention variants, written in CuTe DSL and CUTLASS C++.β534Updated this week
- A project to improve skills of large language modelsβ1,011Updated this week
- Agentic RL on Any Harness at Scaleβ688Updated this week
- (best/better) practices of megatron on veRL and tuning guideβ136May 12, 2026Updated 2 months ago
- Open-source library for scalable, reproducible evaluation of AI models and benchmarks.β314Updated this week
- A lightweight, AI-native training framework for large language models. Designed for fast iteration, reproducible experiments, and modularβ¦β576May 18, 2026Updated 2 months ago
- Ongoing research training transformer models at scaleβ17,125Updated this week
- SkyRL: A Modular Full-stack RL Library for LLMsβ2,081Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- high-performance linear attention kernel library built on TileLangβ597Updated this week
- An efficient implementation of the NSA (Native Sparse Attention) kernelβ133Jun 24, 2025Updated last year
- PyTorch Single Controllerβ1,060Updated this week
- Domain-specific language designed to streamline the development of high-performance GPU/CPU/Accelerators kernelsβ6,674Updated this week
- π₯ A minimal training framework for scaling FLA modelsβ403Apr 22, 2026Updated 2 months ago
- PyTorch-native post-training at scaleβ696Updated this week
- Compact and Agent-Native MoE Training Systemβ290Updated this week