A general large multimodal model for 4D scene understanding
☆15Jul 31, 2025Updated last year
Alternatives and similar repositories for LLaVA-4D
Users that are interested in LLaVA-4D are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICRA 2024]This is the official repo of paper "HR-APR: APR-agnostic Framework with Uncertainty Estimation and Hierarchical Refinement for…☆11Feb 10, 2025Updated last year
- On solutions to the problem of Event Collapse in Motion Compensation frameworks☆15Jan 21, 2023Updated 3 years ago
- [CVPR'26] UniGame code implementation☆20Apr 21, 2026Updated 4 months ago
- DepthART official implementation☆40Oct 28, 2024Updated last year
- smplify code for point cloud based HMR☆10Jan 11, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code release for paper EventCLIP: Adapting CLIP for Event-based Object Recognition☆26Nov 9, 2023Updated 2 years ago
- The official implementation of "Motion-prior Contrast Maximization for Dense Continuous-Time Motion Estimation" (ECCV 2024)☆27Mar 12, 2025Updated last year
- Masked Event Modeling: Self-Supervised Pretraining for Event Cameras (WACV '24)☆21Mar 19, 2024Updated 2 years ago
- Taming Contrast Maximization for Learning Sequential Event-based Optical Flow Estimation☆21Oct 2, 2023Updated 2 years ago
- CoSense3D Documentation☆10Aug 25, 2024Updated 2 years ago
- PyTorch implementation of Hierarchical Neural Memory Network☆50Feb 28, 2024Updated 2 years ago
- ☆12Jul 7, 2023Updated 3 years ago
- This is a project on visual spatial reasoning tasks-SIBench☆28Jan 12, 2026Updated 8 months ago
- [ICCV-2023] The official repository of our paper "TMA: Temporal Motion Aggregation for Event-based Optical Flow".☆35Mar 8, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A PyTorch implementation of Minimal-IK to restore SMPL parameters from skeleton/keypoints + point cloud, with VPoser pose prior for bette…☆23Dec 28, 2023Updated 2 years ago
- [CVPR 2026] Few-Shot Incremental 3D Object Detection in Dynamic Indoor Environments☆18Aug 24, 2026Updated 3 weeks ago
- The repo for code, that hasn't been published yet☆15May 14, 2025Updated last year
- [CVPR 2025] 3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer☆101May 26, 2025Updated last year
- ☆24Mar 22, 2022Updated 4 years ago
- Source code for AEDNet paper from ACMMM 2022☆13May 22, 2026Updated 3 months ago
- Learnable drift compensation (LDC) reduces semantic drift in continual learning using a trainable projector to map between tasks.☆19Nov 13, 2024Updated last year
- DisTime: Distribution-based Time Representation for Video Large Language Models.☆22Jul 10, 2025Updated last year
- [CVPR 2025] FASTer: Focal Token Acquiring-and-Scaling Transformer for Long-term 3D Object Detection☆15Jun 7, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Official implementation of VLAA-GUI series☆37Jun 20, 2026Updated 3 months ago
- [CVPR 2025] The code for paper ''Video-3D LLM: Learning Position-Aware Video Representation for 3D Scene Understanding''.☆224Jun 4, 2025Updated last year
- Seeing World Dynamics in a Nutshell☆114Mar 18, 2025Updated last year
- The official repo for SpaceVista: All-Scale Visual Spatial Reasoning from mm to km.☆44May 26, 2026Updated 3 months ago
- The source code for "LaSe-E2V: Towards Language-guided Semantic-Aware Event-to-Video Reconstruction"☆10Jul 5, 2024Updated 2 years ago
- ☆46Oct 25, 2023Updated 2 years ago
- BetterNet is a state-of-the-art deep learning model for accurate and efficient polyp segmentation in medical images. It combines Efficien…☆14May 8, 2024Updated 2 years ago
- UniDoc-RL: Unified Document Understanding with Reinforcement Learning☆18May 21, 2026Updated 3 months ago
- [ICCV 2025] This is the official implementation of POMATO: Marrying Pointmap Matching with Temporal Motions for Dynamic 3D Reconstruction☆124Aug 9, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Learning Dense and Continuous Optical Flow from an Event Camera (TIP 2022)☆62Jan 13, 2023Updated 3 years ago
- Implementation of Learning without Prejudices: Continual Unbiased Learning via Benign and Malignant Forgetting (ICLR 2023)☆13Apr 14, 2023Updated 3 years ago
- Code for paper "PoseEmbroider:Towards a 3D, Visual, Semantic-aware Human Pose Representation" (ECCV 2024)☆19Nov 18, 2024Updated last year
- [CVPR 2026 Highlight] MonoCoP: Unleashing the Power of Chain-of-Prediction for Monocular 3D Object Detection☆26Aug 3, 2026Updated last month
- ☆16Jan 27, 2026Updated 7 months ago
- [ICCV 2025] Controllable 3D Outdoor Scene Generation via Scene Graphs☆72Jul 13, 2026Updated 2 months ago
- ☆15Dec 15, 2025Updated 9 months ago