Qwen-AgentWorld: Language World Models for General Agents
☆853Jul 20, 2026Updated this week
Alternatives and similar repositories for Qwen-AgentWorld
Users that are interested in Qwen-AgentWorld are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Awesome list and survey website for agents in the era of experience☆159Updated this week
- EdgeBench: Unveiling scaling laws of learning from real-world environments☆365Updated this week
- Post-Trained MoE Can Skip Half Experts via Self-Distillation☆38May 19, 2026Updated 2 months ago
- DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms☆6,702Jul 9, 2026Updated last week
- Scalable pipeline for synthesizing verifiable RLVR training data for computer-use agents☆177May 26, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- "AgentSpace: Human + Agents. One Team. One Workspace"☆706Updated this week
- OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks☆196Jul 9, 2026Updated last week
- Training terminal-agents☆234Updated this week
- CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning☆37Aug 28, 2025Updated 10 months ago
- slime is an LLM post-training framework for RL Scaling.☆7,551Updated this week
- UniRL is a Framework for Unified Multimodal Model Reinforcement Learning☆826Updated this week
- SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning☆344Updated this week
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,575Updated this week
- OpenClaw-RL: Train any agent simply by talking☆5,588May 23, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆175Feb 12, 2026Updated 5 months ago
- Agentic RL on Any Harness at Scale☆688Jul 15, 2026Updated last week
- Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning☆412May 28, 2026Updated last month
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆22,571Updated this week
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆66Apr 13, 2026Updated 3 months ago
- Qwen3.6 is the large language model series developed by Qwen team, Alibaba Group.☆3,703Jun 3, 2026Updated last month
- Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.☆138Apr 2, 2026Updated 3 months ago
- CUA-Gym-Hub: mock web apps as reproducible RL training environments for computer-use agents☆65Jul 9, 2026Updated last week
- Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe☆830Jun 29, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An in-the-wild benchmark for AI agents in the OpenClaw Environment.☆480Updated this week
- Reinforcement Learning via Self-Distillation (SDPO)☆1,017Jul 1, 2026Updated 2 weeks ago
- Code of EMNLP 2025 paper 'UltraIF: Advancing Instruction Following from the Wild'.☆21Apr 3, 2025Updated last year
- Data recipes and robust infrastructure for training AI agents☆260Updated this week
- The roadmap of long-horizon agents☆272Updated this week
- Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond☆286Jun 27, 2026Updated 3 weeks ago
- ☆343Jun 9, 2026Updated last month
- A question-conditioned, reasoning-aware image editor designed to serve as a decoupled visual reasoning assistant for Multimodal Large Lan…☆23May 25, 2026Updated last month
- Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence☆838Jul 10, 2026Updated last week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A generalist autonomous research agent — runs experiments, researches, and iteratively optimizes, autonomously.☆954Jul 12, 2026Updated last week
- Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B☆1,487Jun 17, 2026Updated last month
- Awesome List for On-Policy Distillation☆759Jun 23, 2026Updated 3 weeks ago
- [CVPR 2026] An official implementation of "Think Visually, Reason Textually: Vision-Language Synergy in ARC"☆46Nov 26, 2025Updated 7 months ago
- Official implementation of GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization☆490May 20, 2026Updated 2 months ago
- Unlimited OCR Works: Welcome the Era of One-shot Long-horizon Parsing.☆14,678Jul 3, 2026Updated 2 weeks ago
- [CVPR 2026] Official release of "Spatial-SSRL: Enhancing Spatial Understanding via Self-Supervised Reinforcement Learning"☆133Apr 7, 2026Updated 3 months ago