4D-VLA: Spatiotemporal Vision-Language-Action Pretraining with Cross-Scene Calibration. Accepted to NeurIPS 2025.
☆58Jan 10, 2026Updated 7 months ago
Alternatives and similar repositories for 4D-VLA
Users that are interested in 4D-VLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2026] UniUGG: Unified 3D Understanding and Generation via Geometric-Semantic Encoding☆63Jul 16, 2026Updated 3 weeks ago
- From Flatland to Space (SPAR). Accepted to NeurIPS 2025 Datasets & Benchmarks. A large-scale dataset & benchmark for 3D spatial perceptio…☆93Jan 5, 2026Updated 7 months ago
- [ICRA 2026] Relative Position Matters: Trajectory Prediction and Planning with Polar Representation☆15Feb 5, 2026Updated 6 months ago
- [ICLR 2026] Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining☆31Apr 26, 2026Updated 3 months ago
- Reinforcing Action Policies by Prophesying☆42Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆48Jun 3, 2025Updated last year
- [CVPR 2025] TensoFlow: Tensorial Flow-based Sampler for Inverse Rendering☆15Sep 20, 2025Updated 10 months ago
- [NeurIPS 2025] DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge☆366Jan 6, 2026Updated 7 months ago
- [ICCV2025] BézierGS: Dynamic Urban Scene Reconstruction with Bézier Curve Gaussian Splatting☆139Sep 3, 2025Updated 11 months ago
- Official implementation of PriorVLA.☆17May 11, 2026Updated 3 months ago
- [ICLR 2026] Unified Vision-Language-Action Model☆318Oct 15, 2025Updated 9 months ago
- [ICLR 2025] Diffusion²: Dynamic 3D Content Generation via Score Composition of Video and Multi-view Diffusion Models☆60Mar 18, 2025Updated last year
- [CVPR 2025] VidBot: Learning Generalizable 3D Actions from In-the-Wild 2D Human Videos for Zero-Shot Robotic Manipulation☆51Apr 10, 2026Updated 4 months ago
- [ECCV 2026] Official implementation of CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding.☆52Sep 15, 2025Updated 10 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICLR 2025] GS-LiDAR: Generating Realistic LiDAR Point Clouds with Panoramic Gaussian Splatting☆156Mar 18, 2025Updated last year
- Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model [ICLR2026]☆279Jul 7, 2026Updated last month
- ICCV2025☆172Dec 10, 2025Updated 8 months ago
- [CVPR 2025] Bridging Past and Future: End-to-End Autonomous Driving with Historical Prediction and Planning☆108Apr 7, 2025Updated last year
- [ICCV 2025] Driving Scene Synthesis on Free-form Trajectories with Generative Prior☆41Jun 28, 2025Updated last year
- [CVPR2026] SGDrive: Scene-to-Goal Hierarchical World Cognition for Autonomous Driving☆72Jul 2, 2026Updated last month
- [ICRA2026] ImagiDrive: A Unified Imagination-and-Planning Framework for Autonomous Driving☆24Jul 29, 2026Updated last week
- ☆63Jul 6, 2025Updated last year
- RynnVLA-002: A Unified Vision-Language-Action and World Model☆1,104Dec 2, 2025Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- NORA-1.5: A Vision-Language-Action Model Trained using World Model- and Action-based Preference Rewards☆109Jan 11, 2026Updated 7 months ago
- The offical repo for paper "VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers" (ICCV 2025)☆133Nov 15, 2025Updated 8 months ago
- [CoRL 2025] Pretraining code for FLOWER VLA on OXE☆44Sep 22, 2025Updated 10 months ago
- Official PyTorch implementation for ICML 2025 paper: UP-VLA.☆61Jan 20, 2026Updated 6 months ago
- Code Repository for ControlVLA, CoRL2025.☆100Oct 26, 2025Updated 9 months ago
- [NeurIPS 2025] Future-Aware End-to-End Driving: Bidirectional Modeling of Trajectory Planning and Scene Evolution☆73Feb 4, 2026Updated 6 months ago
- [ICCV 2025] FVP: 4D Visual Pre-training for Robot Learning☆17Sep 5, 2025Updated 11 months ago
- [CVPR'2026] "MM-ACT: Learn from Multimodal Parallel Generation to Act"☆117Mar 13, 2026Updated 4 months ago
- Official code of "Dynamic-eDiTor: Training-Free Text-Driven 4D Scene Editing with Multimodal Diffusion Transformer"☆19May 28, 2026Updated 2 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Learning Visual Feature-Based World Models via Residual Latent Action☆44May 11, 2026Updated 3 months ago
- [CVPR 2026] HiF-VLA: An efficient, bidirectional spatiotemporal expansion Vision-Language-Action Model☆76Mar 11, 2026Updated 5 months ago
- ☆75Jan 8, 2025Updated last year
- [IROS 2025 oral] Official implementation of NOLO: Navigate Only Look Once☆22Nov 13, 2025Updated 8 months ago
- Flowing from Vision to Action: Noise-Free Flow Matching Policy Learning 🎉[ICLR 2026]☆136May 14, 2026Updated 2 months ago
- ☆26Jun 29, 2025Updated last year
- AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation☆39Jul 25, 2025Updated last year