π Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support
β961Sep 19, 2026Updated this week
Alternatives and similar repositories for Automodel
Users that are interested in Automodel are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Training library for Megatron-based models with bidirectional Hugging Face conversion capabilityβ920Updated this week
- Scalable toolkit for efficient model reinforcementβ2,024Updated this week
- A library for exporting models including NeMo and Hugging Face to optimized inference backends, and deploying them for efficient queryingβ42Updated this week
- Scalable data pre processing and curation toolkit for LLMsβ1,770Updated this week
- VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zooβ2,214Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Accelerating MoE with IO and Tile-aware Optimizationsβ769Aug 29, 2026Updated 3 weeks ago
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.β2,946Updated this week
- A tool to configure, launch and manage your machine learning experiments.β258Updated this week
- β274Updated this week
- A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hβ¦β3,544Updated this week
- Best practices for training DeepSeek, Mixtral, Qwen and other MoE models using Megatron Core.β202May 29, 2026Updated 3 months ago
- Bridge Megatron-Core to Hugging Face/Reinforcement Learningβ232Jun 15, 2026Updated 3 months ago
- An agentic-first RL framework for research (9k lines).β1,120Updated this week
- Megatron's multi-modal data loaderβ383Sep 10, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- slime is an LLM post-training framework for RL Scaling.β8,507Updated this week
- A Distributed Attention Towards Linear Scalability for Ultra-Long Context, Heterogeneous Data Trainingβ943Updated this week
- Train speculative decoding models effortlessly and port them smoothly to SGLang serving.β1,178Updated this week
- Evaluate and improve models and agents using environmentsβ1,199Updated this week
- π Efficient implementations for emerging model architecturesβ5,766Updated this week
- A PyTorch native platform for training generative AI modelsβ5,749Updated this week
- A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculativeβ¦β3,828Updated this week
- β37Aug 7, 2025Updated last year
- Distributed Compiler and Optimized Parallel Kernelsβ1,546Updated this week
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Checkpoint-engine is a simple middleware to update model weights in LLM inference enginesβ1,004Updated this week
- CUDA kernels for linear attention variants, written in CuTe DSL and CUTLASS C++.β548Sep 8, 2026Updated last week
- Agentic RL on Any Harness at Scaleβ841Aug 13, 2026Updated last month
- A project to improve skills of large language modelsβ1,041Updated this week
- Minimalistic large language model 3D-parallelism trainingβ2,825Updated this week
- SkyRL: A Modular Full-stack RL Library for LLMsβ2,328Updated this week
- Ongoing research training transformer models at scaleβ17,947Updated this week
- (best/better) practices of megatron on veRL and tuning guideβ138May 12, 2026Updated 4 months ago
- Open-source library for scalable, reproducible evaluation of AI models and benchmarks.β341Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A lightweight, AI-native training framework for large language models. Designed for fast iteration, reproducible experiments, and modularβ¦β587Sep 4, 2026Updated 2 weeks ago
- Domain-specific language designed to streamline the development of high-performance GPU/CPU/Accelerators kernelsβ7,433Updated this week
- high-performance linear attention kernel library built on TileLangβ704Updated this week
- An efficient implementation of the NSA (Native Sparse Attention) kernelβ135Jun 24, 2025Updated last year
- PyTorch Single Controllerβ1,076Updated this week
- A Quirky Assortment of CuTe Kernelsβ1,147Updated this week
- FlashInfer: Kernel Library for LLM Servingβ6,452Updated this week