[CVPR 2026 (Highlight)] 4D-RGPT: Toward Region-level 4D Understanding via Perceptual Distillation
☆39Jun 11, 2026Updated 3 months ago
Alternatives and similar repositories for 4D-RGPT
Users that are interested in 4D-RGPT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code of "Dynamic-eDiTor: Training-Free Text-Driven 4D Scene Editing with Multimodal Diffusion Transformer"☆20May 28, 2026Updated 3 months ago
- (CVPR 2024) "Unsegment Anything by Simulating Deformation"☆29May 27, 2024Updated 2 years ago
- (3DV 2026 Oral) L4P -- a feed-forward foundational model designed for multiple low-level 4D vision perception tasks.☆76Dec 9, 2025Updated 9 months ago
- ☆75Jan 8, 2025Updated last year
- [RA-L'24, IROS'24] Official PyTorch Implementation of "Uni-DVPS: Unified Model for Depth-Aware Video Panoptic Segmentation"☆13Oct 11, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Strefer: Empowering Video LLMs with Space-Time Referring and Reasoning via Synthetic Instruction Data☆19Jun 2, 2026Updated 3 months ago
- [CVPR 2026] HiSpatial: Taming Hierarchical 3D Spatial Understanding in Vision-Language Models☆42Jul 2, 2026Updated 2 months ago
- ☆13Mar 28, 2025Updated last year
- CVPR '26 Highlight☆24May 6, 2026Updated 4 months ago
- Pure C piecewise jerk path optimizer of Apollo for S-L Planning☆12Mar 26, 2024Updated 2 years ago
- Sa2VA-i is an improved version of the popular Sa2VA model☆17Nov 25, 2025Updated 9 months ago
- Evaluation script for RoboSpatial-Home, a benchmark for spatial reasoning in 2D and 3D vision-language models.☆23May 14, 2026Updated 3 months ago
- [NeurIPS 2024] Artemis: Towards Referential Understanding in Complex Videos☆27Apr 8, 2025Updated last year
- [RAL 2026] GSO-SLAM☆23Jun 20, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2026]"Thinking in Dynamics: How Multimodal Large Language Models Perceive, Track, and Reason Dynamics in Physical 4D World"☆17Jul 7, 2026Updated 2 months ago
- Official implementation of No Pose, No Problem in 4D: Feed-Forward Dynamic Gaussians from Unposed Multi-View Videos☆92Sep 6, 2026Updated last week
- Toolkit to transfer the carla format to cosmos RDS-HQ☆15Jun 17, 2025Updated last year
- ☆39Aug 25, 2025Updated last year
- ☆10Dec 12, 2023Updated 2 years ago
- The official code implementation for SynTable - A Synthetic Data Generation Pipeline for Unseen Object Amodal Instance Segmentation of Cl…☆35May 10, 2025Updated last year
- ☆63Jul 3, 2023Updated 3 years ago
- [ICCV 2025] V2XPnP: Vehicle-to-Everything Spatio-Temporal Fusion for Multi-Agent Perception and Prediction☆57Dec 2, 2025Updated 9 months ago
- Localization via embodied dialog on the navigation graph☆15Apr 18, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Package for the teleoperation of UR5+Allegro Robot composite☆12Apr 28, 2022Updated 4 years ago
- code-hosting for Terminus plugin (QGIS)☆12Mar 12, 2022Updated 4 years ago
- [ICIP 2025 Best Student Paper Award] Official code release for "Pose-free 3D Gaussian splatting via shape-ray estimation"☆23Apr 2, 2026Updated 5 months ago
- [NeurIPS 2024] Visual Perception by Large Language Model’s Weights☆56Mar 31, 2025Updated last year
- CVPR 2025' Instruct-4DGS: Efficient Dynamic Scene Editing via 4D Gaussian-based Static-Dynamic Separation☆35Sep 21, 2025Updated 11 months ago
- [ICCV2025] All in One: Visual-Description-Guided Unified Point Cloud Segmentation☆34Jul 25, 2025Updated last year
- Vehicle Trajectory Prediction Library☆16Feb 5, 2024Updated 2 years ago
- ☆29Apr 8, 2025Updated last year
- Official implementation of LoomVideo: Unifying Multimodal Inputs into Video Generation and Editing☆90Jul 17, 2026Updated last month
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Offical Implementation of Captain-Safari [CVPR 2026]☆48Apr 5, 2026Updated 5 months ago
- ☆71Apr 8, 2026Updated 5 months ago
- ☆10Jun 26, 2024Updated 2 years ago
- ☆27May 26, 2026Updated 3 months ago
- [ICML 2024] Sparse Model Inversion: Efficient Inversion of Vision Transformers with Less Hallucination☆14Apr 29, 2025Updated last year
- [ICCV 2025] VLM4D: Towards Spatiotemporal Awareness in Vision Language Models☆56Nov 20, 2025Updated 9 months ago
- Spatial Aptitude Training for Multimodal Langauge Models☆33Feb 8, 2026Updated 7 months ago