4D-VLA: Spatiotemporal Vision-Language-Action Pretraining with Cross-Scene Calibration. Accepted to NeurIPS 2025.
☆59Jan 10, 2026Updated 7 months ago
Alternatives and similar repositories for 4D-VLA
Users that are interested in 4D-VLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2026] UniUGG: Unified 3D Understanding and Generation via Geometric-Semantic Encoding☆63Jul 16, 2026Updated last month
- From Flatland to Space (SPAR). Accepted to NeurIPS 2025 Datasets & Benchmarks. A large-scale dataset & benchmark for 3D spatial perceptio…☆95Jan 5, 2026Updated 7 months ago
- [ICRA 2026] Relative Position Matters: Trajectory Prediction and Planning with Polar Representation☆15Feb 5, 2026Updated 6 months ago
- [ICLR 2026] Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining☆32Apr 26, 2026Updated 4 months ago
- Reinforcing Action Policies by Prophesying☆41Aug 7, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆48Jun 3, 2025Updated last year
- [CVPR 2025] TensoFlow: Tensorial Flow-based Sampler for Inverse Rendering☆15Sep 20, 2025Updated 11 months ago
- [NeurIPS 2025] DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge☆369Jan 6, 2026Updated 7 months ago
- [ICCV2025] BézierGS: Dynamic Urban Scene Reconstruction with Bézier Curve Gaussian Splatting☆139Sep 3, 2025Updated 11 months ago
- Official implementation of PriorVLA.☆17May 11, 2026Updated 3 months ago
- [ICLR 2026] Unified Vision-Language-Action Model☆321Oct 15, 2025Updated 10 months ago
- [ICLR 2025] Diffusion ²: Dynamic 3D Content Generation via Score Composition of Video and Multi-view Diffusion Models☆60Mar 18, 2025Updated last year
- [CVPR 2025] VidBot: Learning Generalizable 3D Actions from In-the-Wild 2D Human Videos for Zero-Shot Robotic Manipulation☆53Apr 10, 2026Updated 4 months ago
- [ECCV 2026] Official implementation of CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding.☆52Sep 15, 2025Updated 11 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICLR 2025] GS-LiDAR: Generating Realistic LiDAR Point Clouds with Panoramic Gaussian Splatting☆157Mar 18, 2025Updated last year
- Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model [ICLR2026]☆281Jul 7, 2026Updated last month
- ICCV2025☆172Dec 10, 2025Updated 8 months ago
- [CVPR 2025] Bridging Past and Future: End-to-End Autonomous Driving with Historical Prediction and Planning☆108Apr 7, 2025Updated last year
- [ICCV 2025] Driving Scene Synthesis on Free-form Trajectories with Generative Prior☆41Jun 28, 2025Updated last year
- [CVPR2026] SGDrive: Scene-to-Goal Hierarchical World Cognition for Autonomous Driving☆72Updated this week
- [ICRA2026] ImagiDrive: A Unified Imagination-and-Planning Framework for Autonomous Driving☆24Jul 29, 2026Updated last month
- ☆63Jul 6, 2025Updated last year
- RynnVLA-002: A Unified Vision-Language-Action and World Model☆1,122Dec 2, 2025Updated 8 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- NORA-1.5: A Vision-Language-Action Model Trained using World Model- and Action-based Preference Rewards☆109Jan 11, 2026Updated 7 months ago
- The offical repo for paper "VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers" (ICCV 2025)☆133Nov 15, 2025Updated 9 months ago
- [CoRL 2025] Pretraining code for FLOWER VLA on OXE☆44Sep 22, 2025Updated 11 months ago
- Official PyTorch implementation for ICML 2025 paper: UP-VLA.☆61Jan 20, 2026Updated 7 months ago
- Code Repository for ControlVLA, CoRL2025.☆99Oct 26, 2025Updated 10 months ago
- [NeurIPS 2025] Future-Aware End-to-End Driving: Bidirectional Modeling of Trajectory Planning and Scene Evolution☆75Feb 4, 2026Updated 6 months ago
- [ICCV 2025] FVP: 4D Visual Pre-training for Robot Learning☆17Sep 5, 2025Updated 11 months ago
- [CVPR'2026] "MM-ACT: Learn from Multimodal Parallel Generation to Act"☆118Mar 13, 2026Updated 5 months ago
- ☆14Jul 6, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official code of "Dynamic-eDiTor: Training-Free Text-Driven 4D Scene Editing with Multimodal Diffusion Transformer"☆19May 28, 2026Updated 3 months ago
- Learning Visual Feature-Based World Models via Residual Latent Action☆47May 11, 2026Updated 3 months ago
- [CVPR 2026] HiF-VLA: An efficient, bidirectional spatiotemporal expansion Vision-Language-Action Model☆77Mar 11, 2026Updated 5 months ago
- ☆75Jan 8, 2025Updated last year
- Flowing from Vision to Action: Noise-Free Flow Matching Policy Learning 🎉[ICLR 2026]☆140May 14, 2026Updated 3 months ago
- ☆26Jun 29, 2025Updated last year
- AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation☆39Jul 25, 2025Updated last year