4D-VLA: Spatiotemporal Vision-Language-Action Pretraining with Cross-Scene Calibration. Accepted to NeurIPS 2025.
☆59Jan 10, 2026Updated 8 months ago
Alternatives and similar repositories for 4D-VLA
Users that are interested in 4D-VLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2026] UniUGG: Unified 3D Understanding and Generation via Geometric-Semantic Encoding☆64Jul 16, 2026Updated 2 months ago
- From Flatland to Space (SPAR). Accepted to NeurIPS 2025 Datasets & Benchmarks. A large-scale dataset & benchmark for 3D spatial perceptio…☆96Jan 5, 2026Updated 8 months ago
- [ICRA 2026] Relative Position Matters: Trajectory Prediction and Planning with Polar Representation☆15Feb 5, 2026Updated 7 months ago
- [ICLR 2026] Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining☆34Apr 26, 2026Updated 4 months ago
- Reinforcing Action Policies by Prophesying☆44Aug 7, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆49Jun 3, 2025Updated last year
- [CVPR 2025] TensoFlow: Tensorial Flow-based Sampler for Inverse Rendering☆15Sep 20, 2025Updated last year
- [NeurIPS 2025] DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge☆372Jan 6, 2026Updated 8 months ago
- [ICCV2025] BézierGS: Dynamic Urban Scene Reconstruction with Bézier Curve Gaussian Splatting☆141Sep 3, 2025Updated last year
- Official implementation of PriorVLA.☆18May 11, 2026Updated 4 months ago
- [ICLR 2026] Unified Vision-Language-Action Model☆322Oct 15, 2025Updated 11 months ago
- [ICLR 2025] Diffusion²: Dynamic 3D Content Generation via Score Composition of Video and Multi-view Diffusion Models☆60Mar 18, 2025Updated last year
- [CVPR 2025] VidBot: Learning Generalizable 3D Actions from In-the-Wild 2D Human Videos for Zero-Shot Robotic Manipulation☆53Apr 10, 2026Updated 5 months ago
- [ECCV 2026] Official implementation of CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding.☆52Sep 15, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ICLR 2025] GS-LiDAR: Generating Realistic LiDAR Point Clouds with Panoramic Gaussian Splatting☆160Mar 18, 2025Updated last year
- Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model [ICLR2026]☆286Jul 7, 2026Updated 2 months ago
- ICCV2025☆172Dec 10, 2025Updated 9 months ago
- [CVPR 2025] Bridging Past and Future: End-to-End Autonomous Driving with Historical Prediction and Planning☆110Apr 7, 2025Updated last year
- [ICCV 2025] Driving Scene Synthesis on Free-form Trajectories with Generative Prior☆41Jun 28, 2025Updated last year
- [CVPR2026] SGDrive: Scene-to-Goal Hierarchical World Cognition for Autonomous Driving☆74Aug 28, 2026Updated 3 weeks ago
- [ICRA2026] ImagiDrive: A Unified Imagination-and-Planning Framework for Autonomous Driving☆24Jul 29, 2026Updated last month
- ☆63Jul 6, 2025Updated last year
- RynnVLA-002: A Unified Vision-Language-Action and World Model☆1,134Dec 2, 2025Updated 9 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- NORA-1.5: A Vision-Language-Action Model Trained using World Model- and Action-based Preference Rewards☆109Jan 11, 2026Updated 8 months ago
- The offical repo for paper "VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers" (ICCV 2025)☆134Nov 15, 2025Updated 10 months ago
- [CoRL 2025] Pretraining code for FLOWER VLA on OXE☆46Sep 22, 2025Updated 11 months ago
- Official PyTorch implementation for ICML 2025 paper: UP-VLA.☆61Jan 20, 2026Updated 8 months ago
- Code Repository for ControlVLA, CoRL2025.☆103Oct 26, 2025Updated 10 months ago
- [NeurIPS 2025] Future-Aware End-to-End Driving: Bidirectional Modeling of Trajectory Planning and Scene Evolution☆78Feb 4, 2026Updated 7 months ago
- [ICCV 2025] FVP: 4D Visual Pre-training for Robot Learning☆17Sep 5, 2025Updated last year
- [CVPR'2026] "MM-ACT: Learn from Multimodal Parallel Generation to Act"☆118Mar 13, 2026Updated 6 months ago
- ☆14Jul 6, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Official code of "Dynamic-eDiTor: Training-Free Text-Driven 4D Scene Editing with Multimodal Diffusion Transformer"☆20May 28, 2026Updated 3 months ago
- Learning Visual Feature-Based World Models via Residual Latent Action☆47May 11, 2026Updated 4 months ago
- [CVPR 2026] HiF-VLA: An efficient, bidirectional spatiotemporal expansion Vision-Language-Action Model☆78Mar 11, 2026Updated 6 months ago
- [IROS 2025 oral] Official implementation of NOLO: Navigate Only Look Once☆22Nov 13, 2025Updated 10 months ago
- ☆75Jan 8, 2025Updated last year
- Flowing from Vision to Action: Noise-Free Flow Matching Policy Learning 🎉[ICLR 2026]☆141May 14, 2026Updated 4 months ago
- ☆26Jun 29, 2025Updated last year