An agentic-first RL framework for research (9k lines).
☆987Aug 29, 2026Updated this week
Alternatives and similar repositories for labs-molt
Users that are interested in labs-molt are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- slime is an LLM post-training framework for RL Scaling.☆8,313Updated this week
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆2,288Updated this week
- Uni-Agent is a framework for training long-horizon agents.☆542Updated this week
- An LLM post-training framework with vLLM for RL Scaling☆443Updated this week
- Agentic RL on Any Harness at Scale☆819Aug 13, 2026Updated 2 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale☆580Updated this week
- Scalable toolkit for efficient model reinforcement☆1,969Updated this week
- Compact and Agent-Native MoE Training System☆345Aug 22, 2026Updated last week
- SkyRL: A Modular Full-stack RL Library for LLMs☆2,210Updated this week
- 🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support☆885Updated this week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,704Updated this week
- An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models☆3,374Updated this week
- A scalable asynchronous reinforcement learning implementation with in-flight weight updates.☆433Aug 5, 2026Updated 3 weeks ago
- Bridge Megatron-Core to Hugging Face/Reinforcement Learning☆230Jun 15, 2026Updated 2 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Checkpoint-engine is a simple middleware to update model weights in LLM inference engines☆1,003Aug 12, 2026Updated 2 weeks ago
- Agentic RL Training at Scale☆1,993Updated this week
- UniRL is a Framework for Unified Multimodal Model Reinforcement Learning☆921Updated this week
- A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.☆4,907May 17, 2026Updated 3 months ago
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆23,204Updated this week
- Democratizing Reinforcement Learning for LLMs☆5,808Updated this week
- My learning notes for ML SYS.☆6,990Aug 19, 2026Updated last week
- mKernel: fast multi-node, multi-GPU fused kernels☆270Updated this week
- VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo☆2,180Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- TokenSpeed is a speed-of-light LLM inference engine.☆2,042Updated this week
- [MLSys 26] 🥇 Solution for Gated Delta Net Track of MLSys 26 Flash infer competition☆36May 22, 2026Updated 3 months ago
- ☆1,309May 20, 2026Updated 3 months ago
- A project implementing various agentic RL based on the Slime post-training framework☆527Apr 11, 2026Updated 4 months ago
- verl Zero-Mismatch Dense/MoE HuggingFace Rollout☆68Aug 5, 2026Updated 3 weeks ago
- Dataflow-Oriented Reinforcement Learning for (Multi-)Agentic LLMs☆102Jul 26, 2026Updated last month
- Implementation for FP8/INT8 Rollout for RL training without performence drop.☆308Nov 7, 2025Updated 9 months ago
- An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Asy…☆9,960Aug 13, 2026Updated 2 weeks ago
- (best/better) practices of megatron on veRL and tuning guide☆138May 12, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Miles-diffusion is an post-training framework for large-scale diffusion model training and production workloads, forked from and co-evolv…☆67Updated this week
- 🚀 Efficient implementations for emerging model architectures☆5,665Updated this week
- Training library for Megatron-based models with bidirectional Hugging Face conversion capability☆890Updated this week
- A lightweight, AI-native training framework for large language models. Designed for fast iteration, reproducible experiments, and modular…☆586May 18, 2026Updated 3 months ago
- Accelerating MoE with IO and Tile-aware Optimizations☆752Updated this week
- FlashKDA: high-performance Kimi Delta Attention kernels☆1,238Jul 30, 2026Updated last month
- Distributed Compiler and Optimized Parallel Kernels☆1,529Aug 12, 2026Updated 2 weeks ago