π Loong: Synthesize Long CoTs at Scale through Verifiers.
β507Aug 7, 2026Updated last week
Alternatives and similar repositories for loong
Users that are interested in loong are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π« CAMEL: The first and the best multi-agent framework. Finding the Scaling Law of Agents. https://www.camel-ai.orgβ17,588Updated this week
- Benchmark and research code for the paper SWEET-RL Training Multi-Turn LLM Agents onCollaborative Reasoning Tasksβ271May 5, 2025Updated last year
- An automated data pipeline scaling RL to pretraining levelsβ76Jun 2, 2026Updated 2 months ago
- RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.β2,770Jul 24, 2026Updated 3 weeks ago
- π¦οΈ CRAB: Cross-environment Agent Benchmark for Multimodal Language Model Agents. https://crab.camel-ai.org/β425Updated this week
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- π» SETA: Scaling Environments for Terminal Agentsβ136Jul 28, 2026Updated 2 weeks ago
- ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning & ReCall: Learning to Reason with Tool Call for LLMs via Reiβ¦β1,428May 16, 2025Updated last year
- Revisiting Mid-training in the Era of Reinforcement Learning Scalingβ188Jul 23, 2025Updated last year
- [EMNLP 2025] Verification Engineering for RL in Instruction Followingβ60Mar 30, 2026Updated 4 months ago
- ποΈ OASIS: Open Agent Social Interaction Simulations with One Million Agents.β5,027Updated this week
- Understanding R1-Zero-Like Training: A Critical Perspectiveβ1,271Aug 27, 2025Updated 11 months ago
- πΎ OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.β671Jan 29, 2026Updated 6 months ago
- Checkpoint-engine is a simple middleware to update model weights in LLM inference enginesβ996Updated this week
- SkyRL: A Modular Full-stack RL Library for LLMsβ2,153Updated this week
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Simple RL training for reasoningβ3,869Dec 23, 2025Updated 7 months ago
- General Reasoner: Advancing LLM Reasoning Across All Domains [NeurIPS25]β230Nov 27, 2025Updated 8 months ago
- Democratizing Reinforcement Learning for LLMsβ5,784Updated this week
- A version of verl to support diverse tool use [TMLR 2026]β1,030Jul 15, 2026Updated last month
- Official Repo for Open-Reasoner-Zeroβ2,097Jun 2, 2025Updated last year
- A unified suite for generating elite reasoning problems and training high-performance LLMs, including pioneering attention-free architectβ¦β132Jan 31, 2026Updated 6 months ago
- Extrapolating RLVR to General Domains without Verifiersβ205Aug 12, 2025Updated last year
- Scaling RL on advanced reasoning modelsβ693Oct 20, 2025Updated 9 months ago
- Official Implementation of Knowledge Flow Promptingβ35Oct 20, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Our library for RL environments + evalsβ4,518Updated this week
- Recipes to train the self-rewarding reasoning LLMs.β231Mar 2, 2025Updated last year
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenariosβ617Jun 12, 2026Updated 2 months ago
- Scalable RL solution for advanced reasoning of language modelsβ1,868Mar 18, 2025Updated last year
- slime is an LLM post-training framework for RL Scaling.β8,027Updated this week
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Frameworkβ22,970Updated this week
- β1,179Jan 10, 2026Updated 7 months ago
- Streamline on-policy/off-policy distillation workflows in a few lines of codeβ109Aug 5, 2026Updated last week
- [ICML 2025] Programming Every Example: Lifting Pre-training Data Quality Like Experts at Scaleβ273Jul 8, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRLβ5,295Nov 13, 2025Updated 9 months ago
- The original Shared Recurrent Memory Transformer implementationβ36Jul 11, 2025Updated last year
- π¦ OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automationβ20,080Updated this week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.β5,667Updated this week
- Reproducible, flexible LLM evaluationsβ391Mar 24, 2026Updated 4 months ago
- OpenTinker is an RL-as-a-Service infrastructure for foundation modelsβ677Mar 21, 2026Updated 4 months ago
- Train your Agent model via our easy and efficient frameworkβ1,778Dec 5, 2025Updated 8 months ago