π Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support
β995Oct 9, 2026Updated this week
Alternatives and similar repositories for Automodel
Users that are interested in Automodel are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Training library for Megatron-based models with bidirectional Hugging Face conversion capabilityβ950Updated this week
- Scalable toolkit for efficient model reinforcementβ2,054Updated this week
- A library for exporting models including NeMo and Hugging Face to optimized inference backends, and deploying them for efficient queryingβ43Updated this week
- Scalable data pre processing and curation toolkit for LLMsβ1,803Updated this week
- State-of-the-art framework for fast, large-scale training and inference of diffusion modelsβ56May 20, 2026Updated 4 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zooβ2,235Updated this week
- Accelerating MoE with IO and Tile-aware Optimizationsβ778Aug 29, 2026Updated last month
- A tool to configure, launch and manage your machine learning experiments.β261Updated this week
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.β3,075Updated this week
- β281Updated this week
- A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hβ¦β3,571Updated this week
- Best practices for training DeepSeek, Mixtral, Qwen and other MoE models using Megatron Core.β204May 29, 2026Updated 4 months ago
- Bridge Megatron-Core to Hugging Face/Reinforcement Learningβ232Jun 15, 2026Updated 3 months ago
- A scalable, agentic-first, and HuggingFace-native RL framework for research (9k lines).β1,173Updated this week
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Megatron's multi-modal data loaderβ385Updated this week
- slime is an LLM post-training framework for RL Scaling.β8,615Updated this week
- A Distributed Attention Towards Linear Scalability for Ultra-Long Context, Heterogeneous Data Trainingβ949Updated this week
- Train speculative decoding models effortlessly and port them smoothly to SGLang serving.β1,205Updated this week
- Evaluate and improve models and agents using environmentsβ1,228Updated this week
- π Efficient implementations for emerging model architecturesβ5,841Updated this week
- A PyTorch native platform for training generative AI modelsβ5,792Updated this week
- β37Aug 7, 2025Updated last year
- A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculativeβ¦β5,251Updated this week
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Distributed Compiler and Optimized Parallel Kernelsβ1,560Sep 18, 2026Updated 3 weeks ago
- Checkpoint-engine is a simple middleware to update model weights in LLM inference enginesβ1,006Sep 15, 2026Updated 3 weeks ago
- CUDA kernels for linear attention variants, written in CuTe DSL and CUTLASS C++.β554Updated this week
- Agentic RL on Any Harness at Scaleβ855Updated this week
- A project to improve skills of large language modelsβ1,048Updated this week
- Minimalistic large language model 3D-parallelism trainingβ2,834Updated this week
- high-performance linear attention kernel library built on TileLangβ722Sep 30, 2026Updated last week
- SkyRL: A Modular Full-stack RL Library for LLMsβ2,397Updated this week
- Ongoing research training transformer models at scaleβ18,096Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- (best/better) practices of megatron on veRL and tuning guideβ138May 12, 2026Updated 4 months ago
- Open-source library for scalable, reproducible evaluation of AI models and benchmarks.β347Oct 1, 2026Updated last week
- A lightweight, AI-native training framework for large language models. Designed for fast iteration, reproducible experiments, and modularβ¦β587Sep 4, 2026Updated last month
- Domain-specific language designed to streamline the development of high-performance GPU/CPU/Accelerators kernelsβ8,536Updated this week
- An efficient implementation of the NSA (Native Sparse Attention) kernelβ135Jun 24, 2025Updated last year
- PyTorch Single Controllerβ1,078Updated this week
- A Quirky Assortment of CuTe Kernelsβ1,166Sep 17, 2026Updated 3 weeks ago