Qwen-AgentWorld: Language World Models for General Agents
☆936Jul 20, 2026Updated 3 weeks ago
Alternatives and similar repositories for Qwen-AgentWorld
Users that are interested in Qwen-AgentWorld are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Awesome list and survey website for agents in the era of experience☆210Updated this week
- EdgeBench: Unveiling scaling laws of learning from real-world environments☆415Jul 17, 2026Updated 3 weeks ago
- Post-Trained MoE Can Skip Half Experts via Self-Distillation☆39May 19, 2026Updated 2 months ago
- DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms☆6,915Jul 9, 2026Updated last month
- Scalable pipeline for synthesizing verifiable RLVR training data for computer-use agents☆185Jul 27, 2026Updated 2 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- "AgentSpace: Human + Agents. One Team. One Workspace"☆922Jul 24, 2026Updated 2 weeks ago
- Training terminal-agents☆274Updated this week
- CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning☆37Aug 28, 2025Updated 11 months ago
- OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks☆237Updated this week
- slime is an LLM post-training framework for RL Scaling.☆7,832Updated this week
- UniRL is a Framework for Unified Multimodal Model Reinforcement Learning☆892Updated this week
- SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning☆360Jul 18, 2026Updated 3 weeks ago
- The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.☆5,653Updated this week
- OpenClaw-RL: Train any agent simply by talking☆5,626May 23, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆182Feb 12, 2026Updated 5 months ago
- Agentic RL on Any Harness at Scale☆760Updated this week
- Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning☆426May 28, 2026Updated 2 months ago
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆22,900Updated this week
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆68Apr 13, 2026Updated 3 months ago
- Qwen3.6 is the large language model series developed by Qwen team, Alibaba Group.☆3,771Jun 3, 2026Updated 2 months ago
- Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.☆142Jul 22, 2026Updated 2 weeks ago
- CUA-Gym-Hub: mock web apps as reproducible RL training environments for computer-use agents☆71Jul 27, 2026Updated 2 weeks ago
- Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe☆911Jun 29, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An in-the-wild benchmark for AI agents in the OpenClaw Environment.☆504Updated this week
- Reinforcement Learning via Self-Distillation (SDPO)☆1,048Jul 1, 2026Updated last month
- Data recipes and robust infrastructure for training AI agents☆278Updated this week
- Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond☆296Jun 27, 2026Updated last month
- ☆356Jun 9, 2026Updated 2 months ago
- A question-conditioned, reasoning-aware image editor designed to serve as a decoupled visual reasoning assistant for Multimodal Large Lan…☆23May 25, 2026Updated 2 months ago
- Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence☆912Updated this week
- [CVPR 2026] An official implementation of "Think Visually, Reason Textually: Vision-Language Synergy in ARC"☆46Nov 26, 2025Updated 8 months ago
- A generalist autonomous research agent — runs experiments, researches, and iteratively optimizes, autonomously.☆997Jul 12, 2026Updated 3 weeks ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Awesome List for On-Policy Distillation☆817Jul 31, 2026Updated last week
- RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.☆2,765Jul 24, 2026Updated 2 weeks ago
- Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B☆1,550Jun 17, 2026Updated last month
- Official implementation of GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization☆496May 20, 2026Updated 2 months ago
- [CVPR 2026] Official release of "Spatial-SSRL: Enhancing Spatial Understanding via Self-Supervised Reinforcement Learning"☆135Apr 7, 2026Updated 4 months ago
- Unlimited OCR Works: Welcome the Era of One-shot Long-horizon Parsing.☆23,304Jul 29, 2026Updated last week
- Official Implementation of "Visual-ERM: Reward Modeling for Visual Equivalence"☆64Mar 23, 2026Updated 4 months ago