Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning
☆429May 28, 2026Updated 2 months ago
Alternatives and similar repositories for agent-world-model
Users that are interested in agent-world-model are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆186Feb 12, 2026Updated 6 months ago
- GenEnv: Difficulty-Aligned Co-Evolution Between LLM Agents and Environment Simulators☆64Dec 23, 2025Updated 7 months ago
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆70Apr 13, 2026Updated 4 months ago
- The official paper for EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL.☆93Jun 5, 2026Updated 2 months ago
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆617Jun 12, 2026Updated 2 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official Implementation of "Simulating Environments with Reasoning Models for Agent Training"☆67Feb 18, 2026Updated 6 months ago
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆19Jun 2, 2026Updated 2 months ago
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in…☆2,224Jun 9, 2026Updated 2 months ago
- RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.☆2,771Jul 24, 2026Updated 3 weeks ago
- Training and evaluating with OpenReward☆33Apr 28, 2026Updated 3 months ago
- [ICLR 2026] The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Execution☆454Updated this week
- SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning☆941May 17, 2026Updated 3 months ago
- OpenClaw-RL: Train any agent simply by talking☆5,638May 23, 2026Updated 2 months ago
- SkyRL: A Modular Full-stack RL Library for LLMs☆2,162Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Reinforcement Learning via Self-Distillation (SDPO)☆1,061Jul 1, 2026Updated last month
- Code for "CREAM: Consistency Regularized Self-Rewarding Language Models", ICLR 2025.☆30Feb 17, 2025Updated last year
- DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use☆29Mar 13, 2026Updated 5 months ago
- Computer Environments Elicit General Agentic Intelligence in LLMs☆241Jul 21, 2026Updated 3 weeks ago
- [ICML 2026] RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments☆231Apr 30, 2026Updated 3 months ago
- slime is an LLM post-training framework for RL Scaling.☆8,093Updated this week
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,108Jul 13, 2026Updated last month
- ☆61Jul 1, 2026Updated last month
- ☆137Mar 31, 2026Updated 4 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Benchmark and research code for the paper SWEET-RL Training Multi-Turn LLM Agents onCollaborative Reasoning Tasks☆271May 5, 2025Updated last year
- An interface library for RL post training with environments.☆2,504Updated this week
- A version of verl to support diverse tool use [TMLR 2026]☆1,031Jul 15, 2026Updated last month
- Agentic RL on Any Harness at Scale☆787Updated this week
- Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.☆145Jul 22, 2026Updated 3 weeks ago
- Dr. MAS is an end-to-end RL training framework for multi-agent LLM systems, supporting the co-training of multiple (heterogeneous) LLMs.☆154Jul 27, 2026Updated 3 weeks ago
- τ-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains☆1,807Updated this week
- ☆32Feb 11, 2026Updated 6 months ago
- Aligning Agentic World Models via Knowledgeable Experience Learning☆39May 15, 2026Updated 3 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Gym-Anything: Turn any Software into an Agent Environment☆276Aug 4, 2026Updated 2 weeks ago
- Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning☆1,613Aug 10, 2026Updated last week
- [ACL 2026 Main] MCP-Flow: Facilitating LLM Agents to Master Real-World, Diverse and Scaling MCP Tools.☆25Apr 8, 2026Updated 4 months ago
- Code for Paper: Training Software Engineering Agents and Verifiers with SWE-Gym [ICML 2025]☆721Jul 29, 2025Updated last year
- Official repo of Toucan: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments☆265Dec 16, 2025Updated 8 months ago
- SkillsBench evaluates how well skills work and how effective agents are at using them.☆1,688Jul 23, 2026Updated 3 weeks ago
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,673Updated this week