π Loong: Synthesize Long CoTs at Scale through Verifiers.
β509Aug 27, 2026Updated last week
Alternatives and similar repositories for loong
Users that are interested in loong are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π« CAMEL: The first and the best multi-agent framework. Finding the Scaling Law of Agents. https://www.camel-ai.orgβ17,670Updated this week
- Benchmark and research code for the paper SWEET-RL Training Multi-Turn LLM Agents onCollaborative Reasoning Tasksβ271May 5, 2025Updated last year
- An automated data pipeline scaling RL to pretraining levelsβ76Jun 2, 2026Updated 3 months ago
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnosticsβ2,789Aug 23, 2026Updated last week
- π¦οΈ CRAB: Cross-environment Agent Benchmark for Multimodal Language Model Agents. https://crab.camel-ai.org/β426Aug 27, 2026Updated last week
- End-to-end encrypted email - Proton Mail β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- π» SETA: Scaling Environments for Terminal Agentsβ145Jul 28, 2026Updated last month
- ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning & ReCall: Learning to Reason with Tool Call for LLMs via Reiβ¦β1,432May 16, 2025Updated last year
- Revisiting Mid-training in the Era of Reinforcement Learning Scalingβ188Jul 23, 2025Updated last year
- [EMNLP 2025] Verification Engineering for RL in Instruction Followingβ60Mar 30, 2026Updated 5 months ago
- ποΈ OASIS: Open Agent Social Interaction Simulations with One Million Agents.β5,101Aug 27, 2026Updated last week
- Understanding R1-Zero-Like Training: A Critical Perspectiveβ1,274Aug 27, 2025Updated last year
- πΎ OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.β669Jan 29, 2026Updated 7 months ago
- Checkpoint-engine is a simple middleware to update model weights in LLM inference enginesβ1,005Aug 12, 2026Updated 3 weeks ago
- Simple RL training for reasoningβ3,873Dec 23, 2025Updated 8 months ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- SkyRL: A Modular Full-stack RL Library for LLMsβ2,247Updated this week
- General Reasoner: Advancing LLM Reasoning Across All Domains [NeurIPS25]β232Nov 27, 2025Updated 9 months ago
- Democratizing Reinforcement Learning for LLMsβ5,814Aug 24, 2026Updated last week
- A version of verl to support diverse tool use [TMLR 2026]β1,037Jul 15, 2026Updated last month
- Official Repo for Open-Reasoner-Zeroβ2,100Jun 2, 2025Updated last year
- A unified suite for generating elite reasoning problems and training high-performance LLMs, including pioneering attention-free architectβ¦β131Jan 31, 2026Updated 7 months ago
- Extrapolating RLVR to General Domains without Verifiersβ205Aug 12, 2025Updated last year
- Scaling RL on advanced reasoning modelsβ696Oct 20, 2025Updated 10 months ago
- Official Implementation of Knowledge Flow Promptingβ35Oct 20, 2025Updated 10 months ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Our library for RL environments + evalsβ4,589Updated this week
- Recipes to train the self-rewarding reasoning LLMs.β230Mar 2, 2025Updated last year
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenariosβ632Jun 12, 2026Updated 2 months ago
- Scalable RL solution for advanced reasoning of language modelsβ1,872Mar 18, 2025Updated last year
- slime is an LLM post-training framework for RL Scaling.β8,381Updated this week
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Frameworkβ23,290Updated this week
- β1,186Jan 10, 2026Updated 7 months ago
- Streamline on-policy/off-policy distillation workflows in a few lines of codeβ109Aug 5, 2026Updated 3 weeks ago
- [ICML 2025] Programming Every Example: Lifting Pre-training Data Quality Like Experts at Scaleβ273Jul 8, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRLβ5,366Nov 13, 2025Updated 9 months ago
- The original Shared Recurrent Memory Transformer implementationβ37Aug 24, 2026Updated last week
- π¦ OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automationβ20,122Aug 27, 2026Updated last week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.β5,723Updated this week
- Reproducible, flexible LLM evaluationsβ394Mar 24, 2026Updated 5 months ago
- OpenTinker is an RL-as-a-Service infrastructure for foundation modelsβ677Mar 21, 2026Updated 5 months ago
- Train your Agent model via our easy and efficient frameworkβ1,780Dec 5, 2025Updated 9 months ago