Sta8is / FUTURISTLinks
[CVPR 2025] Advancing Semantic Future Prediction through Multimodal Visual Sequence Transformers
☆39Updated 3 months ago
Alternatives and similar repositories for FUTURIST
Users that are interested in FUTURIST are comparing it to the libraries listed below
Sorting:
- Official Github Repo for GEM☆99Updated last month
- [ICCV 2025] Stag-1: Towards Realistic 4D Driving Simulation with Video Generation Model☆92Updated last year
- [ICLR 2025] Official code implementation for the paper "X-Drive: Cross-modality Consistent Multi-Sensor Data Synthesis for Driving Scenar…☆61Updated 9 months ago
- [CVPR 2025] GaussTR: Foundation Model-Aligned Gaussian Transformer for Self-Supervised 3D Spatial Understanding☆190Updated 2 months ago
- [NeurIPS 2025] Official Implementation of DINO-Foresight: Looking into the Future with DINO☆136Updated 2 weeks ago
- [ICCV 2025] Embodied 3D Occupancy Prediction for Vision-based Online Scene Understanding☆64Updated 11 months ago
- official code of *DOME: Taming Diffusion Model into High-Fidelity Controllable Occupancy World Model*☆56Updated 11 months ago
- [ICLR 2025] Semi-Supervised Vision-Centric 3D Occupancy World Model for Autonomous Driving☆49Updated 10 months ago
- GASP: Unifying Geometric and Semantic Self-Supervised Pre-training for Autonomous Driving☆27Updated 8 months ago
- [ECCV 2024] TOD3Cap: Towards 3D Dense Captioning in Outdoor Scenes☆129Updated 9 months ago
- ☆103Updated last year
- ☆130Updated last week
- [NeurIPS 2024] DrivingDojo Dataset: Advancing Interactive and Knowledge-Enriched Driving World Model☆82Updated last year
- [ECCV 2024] WoVoGen: World Volume-aware Diffusion for Controllable Multi-camera Driving Scene Generation☆112Updated 10 months ago
- Street-View Image Generation from a Bird’s-Eye View Layout: Official Codebase☆78Updated last year
- [ICCV 2025] GaussRender: Learning 3D Occupancy with Gaussian Rendering (official repository)☆63Updated 5 months ago
- Code for CVPR2025 paper: Generating Multimodal Driving Scenes via Next-Scene Prediction☆96Updated last month
- [ECCV 2024] Monocular Occupancy Prediction for Scalable Indoor Scenes☆65Updated last year
- This is the official project repository for "DriveGEN: Generalized and Robust 3D Detection in Driving via Controllable Text-to-Image Diff…☆35Updated 3 months ago
- ☆48Updated last month
- [ICCV 2025] Language Driven Occupancy Prediction☆33Updated 11 months ago
- [NeurIPS 2025]Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency☆69Updated 2 months ago
- [ICCV 2025] InfiniCube: Unbounded and Controllable Dynamic 3D Driving Scene Generation with World-Guided Video Models☆84Updated 5 months ago
- [ICRA 2025] Official implementation for "TrackOcc: Camera-based 4D Panoptic Occupancy Tracking"☆52Updated 5 months ago
- Unraveling the Effects of Synthetic Data on End-to-End Autonomous Driving☆31Updated 3 weeks ago
- Source code for NeurIPS paper "POP-3D: Open-Vocabulary 3D Occupancy Prediction from Images"☆112Updated 11 months ago
- ☆30Updated last year
- [ECCV 2024] 4D Contrastive Superflows are Dense 3D Representation Learners☆48Updated last week
- ICCV 2025-PerLDiff: Controllable Street View Synthesis Using Perspective-Layout Diffusion Models☆51Updated 4 months ago
- [NeurIPS 2024] DistillNeRF: Perceiving 3D Scenes from Single-Glance Images by Distilling Neural Fields and Foundation Model Features☆36Updated last year