Digital Agents Meet World Models: A Survey
☆53May 8, 2026Updated 3 months ago
Alternatives and similar repositories for awesome-world-models-for-digital-agents
Users that are interested in awesome-world-models-for-digital-agents are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆45Aug 20, 2026Updated last week
- PixelWizard: Towards Efficient High-Fidelity Video Generation at Ultra-Large Spatial Resolutions☆26May 26, 2026Updated 3 months ago
- [ICML 2026] GUIEvalKit: Open-source Evaluation Toolkit for GUI Agents☆25Feb 26, 2026Updated 6 months ago
- [NeurIPS 2025] Implementation of the paper "BTL-UI: Blink-Think-Link Reasoning Model for GUI Agent"☆19Nov 27, 2025Updated 9 months ago
- [ICCV 2025] Implementation of the paper "Q-Frame: Query-aware Frame Selection and Multi-Resolution Adaptation for Video-LLMs"☆82Oct 25, 2025Updated 10 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- State-of-the-art continious audio tokenization☆42Mar 9, 2026Updated 5 months ago
- Generate a complete audio clip with music, intelligible speech, and sound effects from text in one pass.☆46May 27, 2026Updated 3 months ago
- Official Implementation of GLAP - General Language Audio Pretraining☆76May 14, 2026Updated 3 months ago
- PhoneHarness runtime harness for mixed-action phone agents☆47Jun 17, 2026Updated 2 months ago
- Official PyTorch code for Deep Audio-Signal Holistic Embeddings☆204Nov 7, 2025Updated 9 months ago
- end-to-end text to audio scene generation model☆50Jun 16, 2026Updated 2 months ago
- Boosting the Class-Incremental Learning in 3D Point Clouds via Zero-Collection-Cost Basic Shape Pre-Training☆13Nov 30, 2024Updated last year
- DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents☆24Aug 4, 2025Updated last year
- Code for paper OpenWebRL: Online Multi-Turn Reinforcement Learning for Visual Web Agents☆43Aug 17, 2026Updated last week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- This is the code repo for Findings of EMNLP2022 paper: MICO: a multi-alternative contrastive learning framework for commonsense knowledg…☆10Nov 29, 2022Updated 3 years ago
- [TCSVT 2024] Implementation of the paper "SiT-MLP: A Simple MLP with Point-wise Topology Feature Learning for Skeleton-based Action Recog…☆19Apr 10, 2024Updated 2 years ago
- ZeroGUI: Automating Online GUI Learning at Zero Human Cost☆121Jul 17, 2025Updated last year
- Benchmarking Autonomous Mobile Agents in Agent-User Interactive and MCP-Augmented Environments (ACL 2026)☆258Updated this week
- Cross-lingual Visual Pre-training for Multimodal Machine Translation☆18Dec 28, 2021Updated 4 years ago
- ☆17Oct 30, 2023Updated 2 years ago
- [ACL 2026] VGPO: Visually-Guided Policy Optimization for Multimodal Reasoning☆33Apr 14, 2026Updated 4 months ago
- The official code repo of 1.x-Distill, is a stagewise distillation framework for diversity, high-quality and efficient few-step generatio…☆20Apr 10, 2026Updated 4 months ago
- The code for the paper "Efficient Self-Supervised Video Hashing with Selective State Spaces" (AAAI'25).☆24Aug 2, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆16May 22, 2025Updated last year
- ☆18Jun 14, 2023Updated 3 years ago
- The code for the paper "Embracing Collaboration Over Competition: Condensing Multiple Prompts for Visual In-Context Learning" (CVPR'25).☆16Sep 25, 2025Updated 11 months ago
- ☆10Aug 5, 2019Updated 7 years ago
- ☆47Apr 11, 2024Updated 2 years ago
- ☆19Jul 1, 2026Updated last month
- Official code for ACT: Empowering Decision Transformer with Dynamic Programming via Advantage Conditioning (AAAI'24)☆16Feb 10, 2024Updated 2 years ago
- Complexity Based Prompting for Multi-Step Reasoning☆17Mar 10, 2023Updated 3 years ago
- Trust Region Preference Approximation: A simple and stable reinforcement learning algorithm for LLM reasoning☆14Jun 28, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [NeurIPS 2025] UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents☆61Nov 27, 2025Updated 9 months ago
- Direct preference optimization with f-divergences.☆17Nov 3, 2024Updated last year
- ☆17Oct 31, 2023Updated 2 years ago
- ☆19Jan 22, 2024Updated 2 years ago
- Scalable pipeline for synthesizing verifiable RLVR training data for computer-use agents☆193Aug 13, 2026Updated 2 weeks ago
- The reproduct of the paper - Aligner: Achieving Efficient Alignment through Weak-to-Strong Correction☆21May 29, 2024Updated 2 years ago
- [WWW2024 Oral] Harnessing Multi-Role Capabilities of Large Language Models for Open-Domain Question Answering☆15Apr 22, 2025Updated last year