☆537Jul 20, 2026Updated this week
Alternatives and similar repositories for labs-molt
Users that are interested in labs-molt are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- slime is an LLM post-training framework for RL Scaling.☆7,551Updated this week
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆1,759Updated this week
- A unified framework for building, running, and training general agents at scale.☆424Updated this week
- An LLM post-training framework with vLLM for RL Scaling☆378Updated this week
- Agentic RL on Any Harness at Scale☆688Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Scalable toolkit for efficient model reinforcement☆1,835Updated this week
- Compact and Agent-Native MoE Training System☆290Updated this week
- An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale☆509Updated this week
- SkyRL: A Modular Full-stack RL Library for LLMs☆2,081Updated this week
- 🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support☆743Updated this week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,575Updated this week
- An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models☆3,312Updated this week
- Checkpoint-engine is a simple middleware to update model weights in LLM inference engines☆970Jul 4, 2026Updated 2 weeks ago
- A scalable asynchronous reinforcement learning implementation with in-flight weight updates.☆427Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Bridge Megatron-Core to Hugging Face/Reinforcement Learning☆226Jun 15, 2026Updated last month
- Agentic RL Training at Scale☆1,698Updated this week
- UniRL is a Framework for Unified Multimodal Model Reinforcement Learning☆826Updated this week
- A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.☆4,607May 17, 2026Updated 2 months ago
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆22,571Updated this week
- Democratizing Reinforcement Learning for LLMs☆5,708Updated this week
- My learning notes for ML SYS.☆6,753Updated this week
- VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo☆2,097Updated this week
- mKernel: fast multi-node, multi-GPU fused kernels☆251Jun 21, 2026Updated 3 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆1,298May 20, 2026Updated 2 months ago
- TokenSpeed is a speed-of-light LLM inference engine.☆1,638Updated this week
- A project implementing various agentic RL based on the Slime post-training framework☆503Apr 11, 2026Updated 3 months ago
- Implementation for FP8/INT8 Rollout for RL training without performence drop.☆306Nov 7, 2025Updated 8 months ago
- verl Zero-Mismatch Dense/MoE HuggingFace Rollout☆61Updated this week
- (best/better) practices of megatron on veRL and tuning guide☆136May 12, 2026Updated 2 months ago
- 🚀 Efficient implementations for emerging model architectures☆5,379Updated this week
- A lightweight, AI-native training framework for large language models. Designed for fast iteration, reproducible experiments, and modular…☆576May 18, 2026Updated 2 months ago
- [MLSys 26] 🥇 Solution for Gated Delta Net Track of MLSys 26 Flash infer competition☆35May 22, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Distributed Compiler based on Triton for Parallel Systems☆1,494Updated this week
- A version of verl to support diverse tool use [TMLR 2026]☆1,020Updated this week
- Accelerating MoE with IO and Tile-aware Optimizations☆732Jul 4, 2026Updated 2 weeks ago
- FlashKDA: high-performance Kimi Delta Attention kernels☆462May 26, 2026Updated last month
- NexRL is an ultra-loosely-coupled LLM post-training framework.☆114May 13, 2026Updated 2 months ago
- A Gym for Agentic LLMs☆502Jan 21, 2026Updated 5 months ago
- Dataflow-Oriented Reinforcement Learning for (Multi-)Agentic LLMs☆95Updated this week