π Loong: Synthesize Long CoTs at Scale through Verifiers.
β510Sep 20, 2026Updated this week
Alternatives and similar repositories for loong
Users that are interested in loong are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π« CAMEL: The first and the best multi-agent framework. Finding the Scaling Law of Agents. https://www.camel-ai.orgβ17,768Updated this week
- Benchmark and research code for the paper SWEET-RL Training Multi-Turn LLM Agents onCollaborative Reasoning Tasksβ271May 5, 2025Updated last year
- An automated data pipeline scaling RL to pretraining levelsβ76Jun 2, 2026Updated 3 months ago
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnosticsβ2,806Aug 23, 2026Updated last month
- π¦οΈ CRAB: Cross-environment Agent Benchmark for Multimodal Language Model Agents. https://crab.camel-ai.org/β427Updated this week
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- π» SETA: Scaling Environments for Terminal Agentsβ156Jul 28, 2026Updated last month
- ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning & ReCall: Learning to Reason with Tool Call for LLMs via Reiβ¦β1,437May 16, 2025Updated last year
- Revisiting Mid-training in the Era of Reinforcement Learning Scalingβ188Jul 23, 2025Updated last year
- [EMNLP 2025] Verification Engineering for RL in Instruction Followingβ61Mar 30, 2026Updated 5 months ago
- ποΈ OASIS: Open Agent Social Interaction Simulations with One Million Agents.β5,182Updated this week
- Understanding R1-Zero-Like Training: A Critical Perspectiveβ1,278Aug 27, 2025Updated last year
- πΎ OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.β671Jan 29, 2026Updated 7 months ago
- Checkpoint-engine is a simple middleware to update model weights in LLM inference enginesβ1,005Sep 15, 2026Updated last week
- Simple RL training for reasoningβ3,873Dec 23, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- SkyRL: A Modular Full-stack RL Library for LLMsβ2,346Updated this week
- General Reasoner: Advancing LLM Reasoning Across All Domains [NeurIPS25]β232Nov 27, 2025Updated 9 months ago
- Democratizing Reinforcement Learning for LLMsβ5,836Sep 12, 2026Updated last week
- A version of verl to support diverse tool use [TMLR 2026]β1,044Jul 15, 2026Updated 2 months ago
- Official Repo for Open-Reasoner-Zeroβ2,098Jun 2, 2025Updated last year
- Extrapolating RLVR to General Domains without Verifiersβ206Aug 12, 2025Updated last year
- A unified suite for generating elite reasoning problems and training high-performance LLMs, including pioneering attention-free architectβ¦β133Jan 31, 2026Updated 7 months ago
- Scaling RL on advanced reasoning modelsβ695Oct 20, 2025Updated 11 months ago
- Official Implementation of Knowledge Flow Promptingβ35Oct 20, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Our library for RL environments + evalsβ4,650Updated this week
- Recipes to train the self-rewarding reasoning LLMs.β230Mar 2, 2025Updated last year
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenariosβ643Jun 12, 2026Updated 3 months ago
- Scalable RL solution for advanced reasoning of language modelsβ1,873Mar 18, 2025Updated last year
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Frameworkβ23,601Updated this week
- slime is an LLM post-training framework for RL Scaling.β8,527Updated this week
- β1,190Jan 10, 2026Updated 8 months ago
- Streamline on-policy/off-policy distillation workflows in a few lines of codeβ109Aug 5, 2026Updated last month
- [ICML 2025] Programming Every Example: Lifting Pre-training Data Quality Like Experts at Scaleβ272Jul 8, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRLβ5,446Nov 13, 2025Updated 10 months ago
- The original Shared Recurrent Memory Transformer implementationβ39Aug 24, 2026Updated last month
- π¦ OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automationβ20,146Updated this week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.β5,793Updated this week
- Reproducible, flexible LLM evaluationsβ395Mar 24, 2026Updated 6 months ago
- OpenTinker is an RL-as-a-Service infrastructure for foundation modelsβ679Mar 21, 2026Updated 6 months ago
- Train your Agent model via our easy and efficient frameworkβ1,783Dec 5, 2025Updated 9 months ago