π Loong: Synthesize Long CoTs at Scale through Verifiers.
β506Jul 22, 2026Updated this week
Alternatives and similar repositories for loong
Users that are interested in loong are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π« CAMEL: The first and the best multi-agent framework. Finding the Scaling Law of Agents. https://www.camel-ai.orgβ17,490Updated this week
- Benchmark and research code for the paper SWEET-RL Training Multi-Turn LLM Agents onCollaborative Reasoning Tasksβ271May 5, 2025Updated last year
- An automated data pipeline scaling RL to pretraining levelsβ76Jun 2, 2026Updated last month
- RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.β2,756Updated this week
- π» SETA: Scaling Environments for Terminal Agentsβ127Jul 17, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning & ReCall: Learning to Reason with Tool Call for LLMs via Reiβ¦β1,421May 16, 2025Updated last year
- Revisiting Mid-training in the Era of Reinforcement Learning Scalingβ189Jul 23, 2025Updated last year
- [EMNLP 2025] Verification Engineering for RL in Instruction Followingβ57Mar 30, 2026Updated 3 months ago
- Understanding R1-Zero-Like Training: A Critical Perspectiveβ1,268Aug 27, 2025Updated 10 months ago
- πΎ OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.β666Jan 29, 2026Updated 5 months ago
- Checkpoint-engine is a simple middleware to update model weights in LLM inference enginesβ982Jul 4, 2026Updated 3 weeks ago
- SkyRL: A Modular Full-stack RL Library for LLMsβ2,093Updated this week
- Simple RL training for reasoningβ3,870Dec 23, 2025Updated 7 months ago
- General Reasoner: Advancing LLM Reasoning Across All Domains [NeurIPS25]β229Nov 27, 2025Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Democratizing Reinforcement Learning for LLMsβ5,731Updated this week
- A version of verl to support diverse tool use [TMLR 2026]β1,024Jul 15, 2026Updated last week
- Official Repo for Open-Reasoner-Zeroβ2,096Jun 2, 2025Updated last year
- A unified suite for generating elite reasoning problems and training high-performance LLMs, including pioneering attention-free architectβ¦β132Jan 31, 2026Updated 5 months ago
- Extrapolating RLVR to General Domains without Verifiersβ205Aug 12, 2025Updated 11 months ago
- Scaling RL on advanced reasoning modelsβ692Oct 20, 2025Updated 9 months ago
- Official Implementation of Knowledge Flow Promptingβ35Oct 20, 2025Updated 9 months ago
- Our library for RL environments + evalsβ4,400Updated this week
- Recipes to train the self-rewarding reasoning LLMs.β231Mar 2, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenariosβ592Jun 12, 2026Updated last month
- Scalable RL solution for advanced reasoning of language modelsβ1,865Mar 18, 2025Updated last year
- slime is an LLM post-training framework for RL Scaling.β7,629Updated this week
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Frameworkβ22,654Updated this week
- β1,170Jan 10, 2026Updated 6 months ago
- Streamline on-policy/off-policy distillation workflows in a few lines of codeβ107Updated this week
- [ICML 2025] Programming Every Example: Lifting Pre-training Data Quality Like Experts at Scaleβ271Jul 8, 2025Updated last year
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRLβ5,153Nov 13, 2025Updated 8 months ago
- The original Shared Recurrent Memory Transformer implementationβ36Jul 11, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- π¦ OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automationβ20,063Updated this week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.β5,599Updated this week
- Reproducible, flexible LLM evaluationsβ390Mar 24, 2026Updated 4 months ago
- OpenTinker is an RL-as-a-Service infrastructure for foundation modelsβ676Mar 21, 2026Updated 4 months ago
- Train your Agent model via our easy and efficient frameworkβ1,773Dec 5, 2025Updated 7 months ago
- β1,422Sep 12, 2025Updated 10 months ago
- Fully open data curation for reasoning modelsβ2,308Dec 2, 2025Updated 7 months ago