[CVPR 2026 (Highlight)] 4D-RGPT: Toward Region-level 4D Understanding via Perceptual Distillation
☆39Jun 11, 2026Updated last month
Alternatives and similar repositories for 4D-RGPT
Users that are interested in 4D-RGPT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code of "Dynamic-eDiTor: Training-Free Text-Driven 4D Scene Editing with Multimodal Diffusion Transformer"☆19May 28, 2026Updated 2 months ago
- (CVPR 2024) "Unsegment Anything by Simulating Deformation"☆29May 27, 2024Updated 2 years ago
- (3DV 2026 Oral) L4P -- a feed-forward foundational model designed for multiple low-level 4D vision perception tasks.☆73Dec 9, 2025Updated 7 months ago
- ☆75Jan 8, 2025Updated last year
- [RA-L'24, IROS'24] Official PyTorch Implementation of "Uni-DVPS: Unified Model for Depth-Aware Video Panoptic Segmentation"☆13Oct 11, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [CVPR 2026] HiSpatial: Taming Hierarchical 3D Spatial Understanding in Vision-Language Models☆38Jul 2, 2026Updated last month
- Strefer: Empowering Video LLMs with Space-Time Referring and Reasoning via Synthetic Instruction Data☆19Jun 2, 2026Updated 2 months ago
- ☆13Mar 28, 2025Updated last year
- CVPR '26 Highlight☆24May 6, 2026Updated 2 months ago
- Pure C piecewise jerk path optimizer of Apollo for S-L Planning☆12Mar 26, 2024Updated 2 years ago
- Sa2VA-i is an improved version of the popular Sa2VA model☆17Nov 25, 2025Updated 8 months ago
- Evaluation script for RoboSpatial-Home, a benchmark for spatial reasoning in 2D and 3D vision-language models.☆22May 14, 2026Updated 2 months ago
- [NeurIPS 2024] Artemis: Towards Referential Understanding in Complex Videos☆27Apr 8, 2025Updated last year
- [RAL 2026] GSO-SLAM☆20Jun 20, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Spatial-X: Zero-Shot Vision-and-Language Navigation with Spatial Scene Priors☆22Apr 5, 2026Updated 3 months ago
- Official implementation of No Pose, No Problem in 4D: Feed-Forward Dynamic Gaussians from Unposed Multi-View Videos☆74Jun 26, 2026Updated last month
- [CVPR 2026]"Thinking in Dynamics: How Multimodal Large Language Models Perceive, Track, and Reason Dynamics in Physical 4D World"☆17Jul 7, 2026Updated 3 weeks ago
- Prefect integrations with Microsoft Planetary Computer.☆10Jul 15, 2024Updated 2 years ago
- ☆38Aug 25, 2025Updated 11 months ago
- ☆10Dec 12, 2023Updated 2 years ago
- Official code for Latest Object Memory Management for Temporally Consistent Video Instance Segmentation☆34Sep 17, 2025Updated 10 months ago
- [ICCV 2025] V2XPnP: Vehicle-to-Everything Spatio-Temporal Fusion for Multi-Agent Perception and Prediction☆57Dec 2, 2025Updated 8 months ago
- Package for the teleoperation of UR5+Allegro Robot composite☆12Apr 28, 2022Updated 4 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- code-hosting for Terminus plugin (QGIS)☆12Mar 12, 2022Updated 4 years ago
- [NeurIPS 2024] Visual Perception by Large Language Model’s Weights☆56Mar 31, 2025Updated last year
- CVPR 2025' Instruct-4DGS: Efficient Dynamic Scene Editing via 4D Gaussian-based Static-Dynamic Separation☆33Sep 21, 2025Updated 10 months ago
- [ICCV2025] All in One: Visual-Description-Guided Unified Point Cloud Segmentation☆34Jul 25, 2025Updated last year
- Vehicle Trajectory Prediction Library☆16Feb 5, 2024Updated 2 years ago
- ☆29Apr 8, 2025Updated last year
- Official implementation of LoomVideo: Unifying Multimodal Inputs into Video Generation and Editing☆85Jul 17, 2026Updated 2 weeks ago
- ☆69Apr 8, 2026Updated 3 months ago
- [NeurIPS 2025] SAMA: Towards Multi-Turn Referential Grounded Video Chat with Large Language Models.☆17May 26, 2026Updated 2 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆15Apr 26, 2025Updated last year
- ☆33May 13, 2026Updated 2 months ago
- [AAAI 2026 Oral] STRIDE-QA: Visual Question Answering Dataset for Spatiotemporal Reasoning in Urban Driving Scenes☆17Jan 23, 2026Updated 6 months ago
- [ICML 2024] Sparse Model Inversion: Efficient Inversion of Vision Transformers with Less Hallucination☆14Apr 29, 2025Updated last year
- [AAAI 2026 Oral] VQ-Insight: Teaching VLMs for AI-Generated Video Quality Understanding via Progressive Visual Reinforcement Learning☆24Mar 6, 2026Updated 4 months ago
- Information System for Media monitoring and analysis system Project under ПЦФ BR05236839☆10Oct 14, 2023Updated 2 years ago
- [arXiv 2025] SAGE: Training Smart Any-Horizon Agents for Long Video Reasoning with Reinforcement Learning☆70Dec 17, 2025Updated 7 months ago