[CVPR 2026 (Highlight)] 4D-RGPT: Toward Region-level 4D Understanding via Perceptual Distillation
☆39Jun 11, 2026Updated 2 months ago
Alternatives and similar repositories for 4D-RGPT
Users that are interested in 4D-RGPT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆12Mar 1, 2023Updated 3 years ago
- Official code of "Dynamic-eDiTor: Training-Free Text-Driven 4D Scene Editing with Multimodal Diffusion Transformer"☆19May 28, 2026Updated 2 months ago
- OGScene3D: Incremental Open-Vocabulary 3D Gaussian Scene Graph Mapping for Scene Understanding☆17Mar 18, 2026Updated 5 months ago
- (CVPR 2024) "Unsegment Anything by Simulating Deformation"☆29May 27, 2024Updated 2 years ago
- (3DV 2026 Oral) L4P -- a feed-forward foundational model designed for multiple low-level 4D vision perception tasks.☆75Dec 9, 2025Updated 8 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR 2026] HiSpatial: Taming Hierarchical 3D Spatial Understanding in Vision-Language Models☆39Jul 2, 2026Updated last month
- Strefer: Empowering Video LLMs with Space-Time Referring and Reasoning via Synthetic Instruction Data☆19Jun 2, 2026Updated 2 months ago
- ☆13Mar 28, 2025Updated last year
- CVPR '26 Highlight☆24May 6, 2026Updated 3 months ago
- Pure C piecewise jerk path optimizer of Apollo for S-L Planning☆12Mar 26, 2024Updated 2 years ago
- Sa2VA-i is an improved version of the popular Sa2VA model☆17Nov 25, 2025Updated 8 months ago
- [NeurIPS 2024] Artemis: Towards Referential Understanding in Complex Videos☆27Apr 8, 2025Updated last year
- [RAL 2026] GSO-SLAM☆21Jun 20, 2026Updated 2 months ago
- [CVPR 2026]"Thinking in Dynamics: How Multimodal Large Language Models Perceive, Track, and Reason Dynamics in Physical 4D World"☆17Jul 7, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Toolkit to transfer the carla format to cosmos RDS-HQ☆15Jun 17, 2025Updated last year
- Prefect integrations with Microsoft Planetary Computer.☆10Jul 15, 2024Updated 2 years ago
- ☆10Dec 12, 2023Updated 2 years ago
- The official code implementation for SynTable - A Synthetic Data Generation Pipeline for Unseen Object Amodal Instance Segmentation of Cl…☆35May 10, 2025Updated last year
- [ICCV 2025] V2XPnP: Vehicle-to-Everything Spatio-Temporal Fusion for Multi-Agent Perception and Prediction☆57Dec 2, 2025Updated 8 months ago
- Localization via embodied dialog on the navigation graph☆15Apr 18, 2022Updated 4 years ago
- code-hosting for Terminus plugin (QGIS)☆12Mar 12, 2022Updated 4 years ago
- [ICIP 2025 Best Student Paper Award] Official code release for "Pose-free 3D Gaussian splatting via shape-ray estimation"☆22Apr 2, 2026Updated 4 months ago
- [NeurIPS 2024] Visual Perception by Large Language Model’s Weights☆56Mar 31, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- CVPR 2025' Instruct-4DGS: Efficient Dynamic Scene Editing via 4D Gaussian-based Static-Dynamic Separation☆35Sep 21, 2025Updated 11 months ago
- [ICCV2025] All in One: Visual-Description-Guided Unified Point Cloud Segmentation☆34Jul 25, 2025Updated last year
- Vehicle Trajectory Prediction Library☆16Feb 5, 2024Updated 2 years ago
- ☆29Apr 8, 2025Updated last year
- Offical Implementation of Captain-Safari [CVPR 2026]☆48Apr 5, 2026Updated 4 months ago
- [NeurIPS 2025] SAMA: Towards Multi-Turn Referential Grounded Video Chat with Large Language Models.☆18May 26, 2026Updated 2 months ago
- ☆10Jun 26, 2024Updated 2 years ago
- [ICCV 2025] VLM4D: Towards Spatiotemporal Awareness in Vision Language Models☆55Nov 20, 2025Updated 9 months ago
- Spatial Aptitude Training for Multimodal Langauge Models☆33Feb 8, 2026Updated 6 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Using Kolmogorov Arnold Networks (KANs) instead of MLPs in PointNet for Classification and Segmentation of 3D Point Sets☆15Apr 23, 2026Updated 4 months ago
- [arXiv 2025] SAGE: Training Smart Any-Horizon Agents for Long Video Reasoning with Reinforcement Learning☆70Dec 17, 2025Updated 8 months ago
- [CVPR 2025 🔥]A Large Multimodal Model for Pixel-Level Visual Grounding in Videos☆104Apr 14, 2025Updated last year
- NX workspace for running medusa backend, storefront and admin panel with marketplace functionalities☆16Oct 6, 2022Updated 3 years ago
- Sixty years of deforestation and forest fragmentation in Madagascar☆15Nov 25, 2020Updated 5 years ago
- ☆15Jul 25, 2024Updated 2 years ago
- ☆11Apr 5, 2023Updated 3 years ago