Qwen-AgentWorld: Language World Models for General Agents
☆1,013Jul 20, 2026Updated 2 months ago
Alternatives and similar repositories for Qwen-AgentWorld
Users that are interested in Qwen-AgentWorld are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Awesome list and survey website for agents in the era of experience☆265Sep 11, 2026Updated last week
- EdgeBench: Unveiling scaling laws of learning from real-world environments☆450Sep 9, 2026Updated last week
- Post-Trained MoE Can Skip Half Experts via Self-Distillation☆41Sep 6, 2026Updated 2 weeks ago
- DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms☆7,132Jul 9, 2026Updated 2 months ago
- Scalable pipeline for synthesizing verifiable RLVR training data for computer-use agents☆200Aug 13, 2026Updated last month
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Training terminal-agents☆308Sep 12, 2026Updated last week
- "AgentSpace: Human + Agents. One Team. One Workspace"☆995Jul 24, 2026Updated last month
- CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning☆37Aug 28, 2025Updated last year
- OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks☆320Updated this week
- slime is an LLM post-training framework for RL Scaling.☆8,507Updated this week
- SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning☆383Updated this week
- UniRL is a Framework for Unified Multimodal Model Reinforcement Learning☆970Updated this week
- OpenClaw-RL: Train any agent simply by talking☆5,691May 23, 2026Updated 3 months ago
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,776Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆197Sep 3, 2026Updated 2 weeks ago
- Agentic RL on Any Harness at Scale☆841Aug 13, 2026Updated last month
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆23,492Updated this week
- Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning☆456May 28, 2026Updated 3 months ago
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆74Apr 13, 2026Updated 5 months ago
- Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.☆147Jul 22, 2026Updated last month
- Data recipes and robust infrastructure for training AI agents☆295Updated this week
- Qwen3.8 is the large language model series developed by Qwen team, Alibaba Group.☆4,139Aug 17, 2026Updated last month
- An in-the-wild benchmark for AI agents in the production harness.☆523Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe☆1,020Aug 20, 2026Updated last month
- CUA-Gym-Hub: mock web apps as reproducible RL training environments for computer-use agents☆74Sep 3, 2026Updated 2 weeks ago
- Reinforcement Learning via Self-Distillation (SDPO)☆1,100Jul 1, 2026Updated 2 months ago
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics☆2,803Aug 23, 2026Updated 3 weeks ago
- Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond☆315Sep 3, 2026Updated 2 weeks ago
- ☆362Jun 9, 2026Updated 3 months ago
- A question-conditioned, reasoning-aware image editor designed to serve as a decoupled visual reasoning assistant for Multimodal Large Lan…☆23May 25, 2026Updated 3 months ago
- Gym-Anything: Turn any Software into an Agent Environment☆287Updated this week
- Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence☆961Updated this week
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆640Jun 12, 2026Updated 3 months ago
- [CVPR 2026] An official implementation of "Think Visually, Reason Textually: Vision-Language Synergy in ARC"☆47Nov 26, 2025Updated 9 months ago
- A generalist autonomous research agent — runs experiments, researches, and iteratively optimizes, autonomously.☆1,077Sep 8, 2026Updated last week
- Official implementation of GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization☆504May 20, 2026Updated 4 months ago
- Awesome List for On-Policy Distillation☆866Updated this week
- Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B☆1,576Aug 14, 2026Updated last month
- [CVPR 2026] Official release of "Spatial-SSRL: Enhancing Spatial Understanding via Self-Supervised Reinforcement Learning"☆141Apr 7, 2026Updated 5 months ago