【IEEE T-IV】A systematic survey of multi-modal and multi-task visual understanding foundation models for driving scenarios
☆49May 26, 2024Updated 2 years ago
Alternatives and similar repositories for MM-VUFM4DS
Users that are interested in MM-VUFM4DS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2024] This is official implementation of our CVPR 2024 paper "Building a Strong Pre-Training Baseline for Universal 3D Large-Scale …☆17Jun 11, 2024Updated 2 years ago
- A comprehensive survey of forging vision foundation models for autonomous driving, including challenges, methodologies, and opportunities…☆272Jul 1, 2024Updated 2 years ago
- [ICRA 2024] WLST: Weak Labels Guided Self-training for Weakly-supervised Domain Adaptation on 3D Object Detection☆12Feb 6, 2024Updated 2 years ago
- [ECCV 2024] Embodied Understanding of Driving Scenarios☆209Jul 2, 2025Updated last year
- This repository is for CL3D: Unsupervised Domain Adaptation for Cross-LiDAR 3D Detection.☆27Oct 11, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2023] MV-JAR: Masked Voxel Jigsaw and Reconstruction for LiDAR-Based Self-Supervised Pre-Training☆46Jun 4, 2023Updated 3 years ago
- [CVPR 2024 Highlight] Visual Point Cloud Forecasting☆351Jul 2, 2025Updated last year
- [ECCV 2024] WoVoGen: World Volume-aware Diffusion for Controllable Multi-camera Driving Scene Generation☆114Feb 6, 2025Updated last year
- ☆54Jul 23, 2024Updated 2 years ago
- ☆184Dec 25, 2023Updated 2 years ago
- Simulator designed to generate diverse driving scenarios.☆44Feb 27, 2025Updated last year
- ☆15Feb 27, 2025Updated last year
- Fine-Grained Evaluation of Large Vision-Language Models in Autonomous Driving (ICCV 2025)☆38Mar 26, 2026Updated 4 months ago
- ☆14Jul 9, 2021Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- BEV-LGKD: A Unified LiDAR-Guided Knowledge Distillation Framework for Multi-View BEV 3D Object Detection☆13Apr 2, 2024Updated 2 years ago
- [Official] [IROS 2024] A goal-oriented planning to lift VLN performance for Closed-Loop Navigation: Simple, Yet Effective☆28Apr 4, 2024Updated 2 years ago
- [CVPR 2024] A world model for autonomous driving.☆438Dec 7, 2023Updated 2 years ago
- DetMatch: Two Teachers are Better Than One for Joint 2D and 3D Semi-Supervised Object Detection☆36Jun 9, 2023Updated 3 years ago
- [CVPR 2023] Pytorch implementation of ThinkTwice, a SOTA Decoder for End-to-end Autonomous Driving under BEV.☆254Jul 2, 2025Updated last year
- ☆54Dec 4, 2024Updated last year
- Official PyTorch implementation of CODA-LM(https://arxiv.org/abs/2404.10595)☆103Dec 5, 2024Updated last year
- [ECCV 2024] TOD3Cap: Towards 3D Dense Captioning in Outdoor Scenes☆132Mar 1, 2025Updated last year
- This is the repository for DDS3D(ICRA2023)☆19Nov 10, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICCVW 2025] Simplifying Traffic Anomaly Detection with Video Foundation Models☆18Dec 4, 2025Updated 7 months ago
- [CVPR2024] Multiagent Multitraversal Multimodal Self-Driving: Open MARS Dataset☆63Jun 25, 2024Updated 2 years ago
- BEVContrast: Self-Supervision in BEV Space for Automotive Lidar Point Clouds - Official PyTorch implementation☆81Jun 4, 2024Updated 2 years ago
- Track 3: Sensor Placement☆19Aug 22, 2025Updated 11 months ago
- A curated list of awesome knowledge-driven autonomous driving (continually updated)☆501Jun 7, 2024Updated 2 years ago
- Implementation of "Unsupervised Domain Adaptive 3D Detection with Multi-Level Consistency"☆48Dec 2, 2021Updated 4 years ago
- Adaptive Multimodal Learning for Remote Sensing Data Fusion☆13Dec 22, 2024Updated last year
- Repo for 'VLM-Auto: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes'☆27Oct 10, 2024Updated last year
- [WACV 2024 Survey Paper] Multimodal Large Language Models for Autonomous Driving☆313Mar 14, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Multiple Lidar preprocessor for BEVfusion☆11Aug 25, 2023Updated 2 years ago
- 3D cascade RCNN for object detection on point cloud☆32Aug 13, 2022Updated 3 years ago
- Official Code Release of "FusionAD"☆165Jul 9, 2024Updated 2 years ago
- This repository contains the PyTorch implementation of the ECCV'2022 paper, ProposalContrast: Unsupervised Pre-training for LiDAR-based 3…☆58Sep 7, 2022Updated 3 years ago
- [WACV 2024] MACP: Efficient Model Adaptation for Cooperative Perception.☆19May 3, 2024Updated 2 years ago
- ☆25Jun 17, 2022Updated 4 years ago
- [ICML 2024] CrossGET: Cross-Guided Ensemble of Tokens for Accelerating Vision-Language Transformers☆34Dec 30, 2024Updated last year