Qwen-AgentWorld: Language World Models for General Agents
☆1,028Jul 20, 2026Updated 2 months ago
Alternatives and similar repositories for Qwen-AgentWorld
Users that are interested in Qwen-AgentWorld are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Awesome list and survey website for agents in the era of experience☆277Updated this week
- EdgeBench: Unveiling scaling laws of learning from real-world environments☆465Sep 20, 2026Updated 2 weeks ago
- Post-Trained MoE Can Skip Half Experts via Self-Distillation☆42Sep 6, 2026Updated last month
- DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms☆7,198Jul 9, 2026Updated 3 months ago
- Scalable pipeline for synthesizing verifiable RLVR training data for computer-use agents☆208Aug 13, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Training terminal-agents☆318Sep 20, 2026Updated 2 weeks ago
- "AgentSpace: Human + Agents. One Team. One Workspace"☆1,007Jul 24, 2026Updated 2 months ago
- CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning☆37Aug 28, 2025Updated last year
- OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks☆357Oct 1, 2026Updated last week
- slime is an LLM post-training framework for RL Scaling.☆8,615Updated this week
- OpenClaw-RL: Train any agent simply by talking☆5,717May 23, 2026Updated 4 months ago
- UniRL is a Framework for Unified Multimodal Model Reinforcement Learning☆1,009Updated this week
- [NeurIPS 2026] SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning☆433Sep 14, 2026Updated 3 weeks ago
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,823Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆199Sep 3, 2026Updated last month
- Agentic RL on Any Harness at Scale☆855Updated this week
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆23,805Updated this week
- Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning☆467May 28, 2026Updated 4 months ago
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆78Apr 13, 2026Updated 5 months ago
- Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.☆151Oct 1, 2026Updated last week
- Qwen3.8 is the large language model series developed by Qwen team, Alibaba Group.☆4,251Aug 17, 2026Updated last month
- An in-the-wild benchmark for AI agents in the production harness.☆529Sep 18, 2026Updated 3 weeks ago
- Data recipes and robust infrastructure for training AI agents☆301Sep 28, 2026Updated last week
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe☆1,038Aug 20, 2026Updated last month
- CUA-Gym-Hub: mock web apps as reproducible RL training environments for computer-use agents☆81Sep 3, 2026Updated last month
- Reinforcement Learning via Self-Distillation (SDPO)☆1,115Jul 1, 2026Updated 3 months ago
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics☆2,818Aug 23, 2026Updated last month
- Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond☆325Sep 30, 2026Updated last week
- ☆366Jun 9, 2026Updated 4 months ago
- A question-conditioned, reasoning-aware image editor designed to serve as a decoupled visual reasoning assistant for Multimodal Large Lan…☆23May 25, 2026Updated 4 months ago
- Gym-Anything: Turn any Software into an Agent Environment☆292Sep 26, 2026Updated 2 weeks ago
- Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence☆966Sep 17, 2026Updated 3 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆647Jun 12, 2026Updated 3 months ago
- [CVPR 2026] An official implementation of "Think Visually, Reason Textually: Vision-Language Synergy in ARC"☆47Nov 26, 2025Updated 10 months ago
- A generalist autonomous research agent — runs experiments, researches, and iteratively optimizes, autonomously.☆1,105Sep 8, 2026Updated last month
- Official implementation of GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization☆511May 20, 2026Updated 4 months ago
- Awesome List for On-Policy Distillation☆873Oct 1, 2026Updated last week
- Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B☆1,581Aug 14, 2026Updated last month
- [CVPR 2026] Official release of "Spatial-SSRL: Enhancing Spatial Understanding via Self-Supervised Reinforcement Learning"☆142Apr 7, 2026Updated 6 months ago