Qwen-AgentWorld: Language World Models for General Agents
☆981Jul 20, 2026Updated last month
Alternatives and similar repositories for Qwen-AgentWorld
Users that are interested in Qwen-AgentWorld are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Awesome list and survey website for agents in the era of experience☆244Aug 15, 2026Updated 2 weeks ago
- EdgeBench: Unveiling scaling laws of learning from real-world environments☆432Jul 17, 2026Updated last month
- Post-Trained MoE Can Skip Half Experts via Self-Distillation☆40May 19, 2026Updated 3 months ago
- DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms☆7,056Jul 9, 2026Updated last month
- Scalable pipeline for synthesizing verifiable RLVR training data for computer-use agents☆193Aug 13, 2026Updated 2 weeks ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- "AgentSpace: Human + Agents. One Team. One Workspace"☆955Jul 24, 2026Updated last month
- Training terminal-agents☆289Updated this week
- CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning☆37Aug 28, 2025Updated last year
- OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks☆264Updated this week
- slime is an LLM post-training framework for RL Scaling.☆8,313Updated this week
- UniRL is a Framework for Unified Multimodal Model Reinforcement Learning☆921Updated this week
- SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning☆371Jul 18, 2026Updated last month
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,704Updated this week
- OpenClaw-RL: Train any agent simply by talking☆5,660May 23, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆191Feb 12, 2026Updated 6 months ago
- Agentic RL on Any Harness at Scale☆819Aug 13, 2026Updated 2 weeks ago
- Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning☆438May 28, 2026Updated 3 months ago
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆23,204Updated this week
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆71Apr 13, 2026Updated 4 months ago
- Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.☆146Jul 22, 2026Updated last month
- Qwen3.8 is the large language model series developed by Qwen team, Alibaba Group.☆4,020Aug 17, 2026Updated 2 weeks ago
- CUA-Gym-Hub: mock web apps as reproducible RL training environments for computer-use agents☆71Aug 23, 2026Updated last week
- Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe☆972Aug 20, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- An in-the-wild benchmark for AI agents in the production harness.☆515Aug 17, 2026Updated 2 weeks ago
- Reinforcement Learning via Self-Distillation (SDPO)☆1,081Jul 1, 2026Updated last month
- Data recipes and robust infrastructure for training AI agents☆286Updated this week
- Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond☆302Updated this week
- ☆360Jun 9, 2026Updated 2 months ago
- A question-conditioned, reasoning-aware image editor designed to serve as a decoupled visual reasoning assistant for Multimodal Large Lan…☆23May 25, 2026Updated 3 months ago
- Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence☆938Aug 5, 2026Updated 3 weeks ago
- [CVPR 2026] An official implementation of "Think Visually, Reason Textually: Vision-Language Synergy in ARC"☆47Nov 26, 2025Updated 9 months ago
- A generalist autonomous research agent — runs experiments, researches, and iteratively optimizes, autonomously.☆1,046Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Awesome List for On-Policy Distillation☆846Updated this week
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics☆2,781Aug 23, 2026Updated last week
- Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B☆1,564Aug 14, 2026Updated 2 weeks ago
- Official implementation of GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization☆500May 20, 2026Updated 3 months ago
- [CVPR 2026] Official release of "Spatial-SSRL: Enhancing Spatial Understanding via Self-Supervised Reinforcement Learning"☆138Apr 7, 2026Updated 4 months ago
- Unlimited OCR Works: Welcome the Era of One-shot Long-horizon Parsing.☆24,770Jul 29, 2026Updated last month
- Official Implementation of "Visual-ERM: Reward Modeling for Visual Equivalence"☆65Mar 23, 2026Updated 5 months ago