Learning Visual Feature-Based World Models via Residual Latent Action
☆46May 11, 2026Updated 3 months ago
Alternatives and similar repositories for rla-wm
Users that are interested in rla-wm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Accept by RSS 2026☆229Jun 1, 2026Updated 2 months ago
- code for Imagination-Policy☆16Dec 1, 2024Updated last year
- Self-CorrectingVLA:OnlineActionRefinementviaSparseWorldImagination☆27Apr 14, 2026Updated 4 months ago
- [ICLR 2026] Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining☆32Apr 26, 2026Updated 3 months ago
- This repository contains the code for the paper - "Aligning Text, Images, and 3D Structure Token-by-Token" (CVPR 2026)☆49Jun 11, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official implementation of Paperify. for running the Bilibili account "具身人机", focusing on embodied intelligence, human-computer interacti…☆16Jun 30, 2026Updated last month
- Afford-VLA: Action-Aligned Visual Planning via Internalized Affordance☆21Jul 5, 2026Updated last month
- [CVPR2026] Chain of World: World Model Thinking in Latent Motion☆66Mar 4, 2026Updated 5 months ago
- Official implementation of AMPLIFY: Actionless Motion Priors for Robot Learning from Videos☆53Apr 13, 2026Updated 4 months ago
- [ICRA 2024] WLST: Weak Labels Guided Self-training for Weakly-supervised Domain Adaptation on 3D Object Detection☆12Feb 6, 2024Updated 2 years ago
- ReSemAct: Advancing Fine-Grained Robotic Manipulation via Semantic Structuring and Affordance Refinement☆17Jan 5, 2026Updated 7 months ago
- ☆28Aug 6, 2024Updated 2 years ago
- [RA-L 2025] RT-GuIDE: Real-Time Gaussian Splatting for Information-Driven Exploration☆25Nov 30, 2025Updated 8 months ago
- [CVPR 2026 Highlight] A Frame is Worth One Token: Efficient Generative World Modeling with Delta Tokens☆238Jul 17, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆63Jul 6, 2025Updated last year
- World Modeling by Forecasting Vision Foundation Model Features☆56Jul 25, 2026Updated 3 weeks ago
- [ICRA 2026] History-Aware Visuomotor Policy Learning via Point Tracking☆27Jan 10, 2026Updated 7 months ago
- VLA-RFT: Vision-Language-Action Models with Reinforcement Fine-Tuning☆160Oct 6, 2025Updated 10 months ago
- LUMOS: Language-Conditioned Imitation Learning with World Models☆20Apr 1, 2026Updated 4 months ago
- ☆145Mar 31, 2026Updated 4 months ago
- ☆13Oct 29, 2023Updated 2 years ago
- Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow☆46Mar 28, 2026Updated 4 months ago
- 多足机器人MPC控制器+仿真环境☆13Feb 15, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Zero-WAM, an in-context world model for zero-shot robotic task generalization☆35Jul 8, 2026Updated last month
- 4D-VLA: Spatiotemporal Vision-Language-Action Pretraining with Cross-Scene Calibration. Accepted to NeurIPS 2025.☆58Jan 10, 2026Updated 7 months ago
- The repository provides code for EgoMAN model and dataset creation scripts.☆32Dec 31, 2025Updated 7 months ago
- ☆19Oct 18, 2025Updated 10 months ago
- [AAAI 2026] SmartSplat: Feature-Smart Gaussians for Scalable Compression of Ultra-High-Resolution Images☆20Dec 26, 2025Updated 7 months ago
- A conda-smithy repository for colmap.☆15Aug 14, 2026Updated last week
- [IROS 2025] GraspMAS: Zero-Shot Language-driven Grasp Detection with Multi-Agent System☆18Oct 26, 2025Updated 9 months ago
- [CVPR 2025]🌷This is the official repository of "Two by Two✌️ : Learning Multi-Task Pairwise Objects Assembly for Generalizable Robot Man…☆63Sep 29, 2025Updated 10 months ago
- Robo3R: Enhancing Robotic Manipulation with Accurate Feed-Forward 3D Reconstruction☆73Jun 21, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official PyTorch Implementation of Unified Video Action Model (RSS 2025)☆405Updated this week
- Fast per-point embedding kernel for scene flow estimation☆12Mar 12, 2024Updated 2 years ago
- ☆27Jul 2, 2026Updated last month
- Source code For AAAI 2026 paper: "RaLiFlow: Scene Flow Estimation with 4D Radar and LiDAR Point Clouds"☆15Jun 5, 2026Updated 2 months ago
- This is the official code repo for GLOVER and GLOVER++.☆58Aug 6, 2025Updated last year
- Steering Video World Model Imaginations into High-impact, Plausible Outcomes with Initial Noise Optimization☆19Jun 2, 2026Updated 2 months ago
- [ACM MM'26] MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation☆65May 14, 2026Updated 3 months ago