Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning
☆459May 28, 2026Updated 3 months ago
Alternatives and similar repositories for agent-world-model
Users that are interested in agent-world-model are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆198Sep 3, 2026Updated 3 weeks ago
- [NeurIPS 2026] GenEnv: Difficulty-Aligned Co-Evolution Between LLM Agents and Environment Simulators☆67Updated this week
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆75Apr 13, 2026Updated 5 months ago
- The official paper for EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL.☆96Aug 28, 2026Updated last month
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆645Jun 12, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official Implementation of "Simulating Environments with Reasoning Models for Agent Training"☆68Feb 18, 2026Updated 7 months ago
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆20Jun 2, 2026Updated 3 months ago
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in…☆2,342Jun 9, 2026Updated 3 months ago
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics☆2,806Aug 23, 2026Updated last month
- Training and evaluating with OpenReward☆33Apr 28, 2026Updated 4 months ago
- [ICLR 2026] The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Execution☆489Aug 18, 2026Updated last month
- SkyRL: A Modular Full-stack RL Library for LLMs☆2,353Updated this week
- OpenClaw-RL: Train any agent simply by talking☆5,705May 23, 2026Updated 4 months ago
- [NeurIPS'26] SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning☆983Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Reinforcement Learning via Self-Distillation (SDPO)☆1,104Jul 1, 2026Updated 2 months ago
- Code for "CREAM: Consistency Regularized Self-Rewarding Language Models", ICLR 2025.☆30Feb 17, 2025Updated last year
- DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use☆30Mar 13, 2026Updated 6 months ago
- Computer Environments Elicit General Agentic Intelligence in LLMs☆244Aug 22, 2026Updated last month
- A version of verl to support diverse tool use [TMLR 2026]☆1,044Jul 15, 2026Updated 2 months ago
- [ICML 2026] RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments☆235Apr 30, 2026Updated 4 months ago
- slime is an LLM post-training framework for RL Scaling.☆8,547Updated this week
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,126Sep 12, 2026Updated 2 weeks ago
- Agentic RL on Any Harness at Scale☆843Aug 13, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Benchmark and research code for the paper SWEET-RL Training Multi-Turn LLM Agents onCollaborative Reasoning Tasks☆271May 5, 2025Updated last year
- ☆68Jul 1, 2026Updated 2 months ago
- ☆148Mar 31, 2026Updated 5 months ago
- An interface library for RL post training with environments.☆2,619Updated this week
- Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.☆149Jul 22, 2026Updated 2 months ago
- ☆32Feb 11, 2026Updated 7 months ago
- Dr. MAS is an end-to-end RL training framework for multi-agent LLM systems, supporting the co-training of multiple (heterogeneous) LLMs.☆165Updated this week
- τ-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains☆2,110Sep 19, 2026Updated last week
- [EMNLP 2026] Aligning Agentic World Models via Knowledgeable Experience Learning☆41May 15, 2026Updated 4 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Gym-Anything: Turn any Software into an Agent Environment☆288Updated this week
- Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning☆1,677Updated this week
- [ACL 2026 Main] MCP-Flow: Facilitating LLM Agents to Master Real-World, Diverse and Scaling MCP Tools.☆25Apr 8, 2026Updated 5 months ago
- Code for Paper: Training Software Engineering Agents and Verifiers with SWE-Gym [ICML 2025]☆746Jul 29, 2025Updated last year
- Official repo of Toucan: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments☆268Dec 16, 2025Updated 9 months ago
- Yet another dynamic batch sampler for variable sequence data in PyTorch.☆13Dec 9, 2021Updated 4 years ago
- SkillsBench evaluates how well skills work and how effective agents are at using them.☆1,816Jul 23, 2026Updated 2 months ago