Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning
☆418May 28, 2026Updated 2 months ago
Alternatives and similar repositories for agent-world-model
Users that are interested in agent-world-model are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆179Feb 12, 2026Updated 5 months ago
- GenEnv: Difficulty-Aligned Co-Evolution Between LLM Agents and Environment Simulators☆62Dec 23, 2025Updated 7 months ago
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆66Apr 13, 2026Updated 3 months ago
- The official paper for EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL.☆85Jun 5, 2026Updated last month
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆595Jun 12, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official Implementation of "Simulating Environments with Reasoning Models for Agent Training"☆65Feb 18, 2026Updated 5 months ago
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆18Jun 2, 2026Updated last month
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in…☆2,158Jun 9, 2026Updated last month
- RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.☆2,757Updated this week
- Training and evaluating with OpenReward☆33Apr 28, 2026Updated 3 months ago
- [ICLR 2026] The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Execution☆441Updated this week
- SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning☆909May 17, 2026Updated 2 months ago
- OpenClaw-RL: Train any agent simply by talking☆5,611May 23, 2026Updated 2 months ago
- SkyRL: A Modular Full-stack RL Library for LLMs☆2,102Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Reinforcement Learning via Self-Distillation (SDPO)☆1,027Jul 1, 2026Updated 3 weeks ago
- Code for "CREAM: Consistency Regularized Self-Rewarding Language Models", ICLR 2025.☆29Feb 17, 2025Updated last year
- DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use☆28Mar 13, 2026Updated 4 months ago
- Computer Environments Elicit General Agentic Intelligence in LLMs☆239Jul 21, 2026Updated last week
- [ICML 2026] RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments☆226Apr 30, 2026Updated 2 months ago
- slime is an LLM post-training framework for RL Scaling.☆7,679Updated this week
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,093Jul 13, 2026Updated 2 weeks ago
- ☆58Jul 1, 2026Updated 3 weeks ago
- ☆134Mar 31, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Benchmark and research code for the paper SWEET-RL Training Multi-Turn LLM Agents onCollaborative Reasoning Tasks☆271May 5, 2025Updated last year
- An interface library for RL post training with environments.☆2,452Updated this week
- Dr. MAS is an end-to-end RL training framework for multi-agent LLM systems, supporting the co-training of multiple (heterogeneous) LLMs.☆145Updated this week
- A version of verl to support diverse tool use [TMLR 2026]☆1,026Jul 15, 2026Updated 2 weeks ago
- Agentic RL on Any Harness at Scale☆714Jul 15, 2026Updated 2 weeks ago
- Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.☆140Jul 22, 2026Updated last week
- τ-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains☆1,685Updated this week
- ☆31Feb 11, 2026Updated 5 months ago
- Aligning Agentic World Models via Knowledgeable Experience Learning☆37May 15, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Gym-Anything: Turn any Software into an Agent Environment☆263Updated this week
- Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning☆1,574Updated this week
- [ACL 2026 Main] MCP-Flow: Facilitating LLM Agents to Master Real-World, Diverse and Scaling MCP Tools.☆25Apr 8, 2026Updated 3 months ago
- Code for Paper: Training Software Engineering Agents and Verifiers with SWE-Gym [ICML 2025]☆712Jul 29, 2025Updated last year
- Official repo of Toucan: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments☆260Dec 16, 2025Updated 7 months ago
- SkillsBench evaluates how well skills work and how effective agents are at using them.☆1,596Updated this week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,612Updated this week