A general large multimodal model for 4D scene understanding
☆16Jul 31, 2025Updated 11 months ago
Alternatives and similar repositories for LLaVA-4D
Users that are interested in LLaVA-4D are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICRA 2024]This is the official repo of paper "HR-APR: APR-agnostic Framework with Uncertainty Estimation and Hierarchical Refinement for…☆11Feb 10, 2025Updated last year
- On solutions to the problem of Event Collapse in Motion Compensation frameworks☆15Jan 21, 2023Updated 3 years ago
- [CVPR'26] UniGame code implementation☆20Apr 21, 2026Updated 2 months ago
- Taming Contrast Maximization for Learning Sequential Event-based Optical Flow Estimation☆21Oct 2, 2023Updated 2 years ago
- DepthART official implementation☆24Oct 28, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- PyTorch implementation of E-Motion: Future Motion Simulation via Event Sequence Diffusion☆17Jul 5, 2025Updated last year
- ☆27Nov 15, 2023Updated 2 years ago
- Code release for paper EventCLIP: Adapting CLIP for Event-based Object Recognition☆25Nov 9, 2023Updated 2 years ago
- The official implementation of "Motion-prior Contrast Maximization for Dense Continuous-Time Motion Estimation" (ECCV 2024)☆27Mar 12, 2025Updated last year
- CoSense3D Documentation☆11Aug 25, 2024Updated last year
- PyTorch implementation of Hierarchical Neural Memory Network☆49Feb 28, 2024Updated 2 years ago
- This is a project on visual spatial reasoning tasks-SIBench☆27Jan 12, 2026Updated 6 months ago
- ☆12Jul 7, 2023Updated 3 years ago
- A PyTorch implementation of Minimal-IK to restore SMPL parameters from skeleton/keypoints + point cloud, with VPoser pose prior for bette…☆22Dec 28, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICCV-2023] The official repository of our paper "TMA: Temporal Motion Aggregation for Event-based Optical Flow".☆33Mar 8, 2024Updated 2 years ago
- Responsible Robotic Manipulation☆16Aug 31, 2025Updated 10 months ago
- [CVPR 2026] Few-Shot Incremental 3D Object Detection in Dynamic Indoor Environments☆16Apr 10, 2026Updated 3 months ago
- The repo for code, that hasn't been published yet☆14May 14, 2025Updated last year
- [CVPR 2025] 3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer☆100May 26, 2025Updated last year
- ☆23Mar 22, 2022Updated 4 years ago
- Learnable drift compensation (LDC) reduces semantic drift in continual learning using a trainable projector to map between tasks.☆19Nov 13, 2024Updated last year
- Source code for AEDNet paper from ACMMM 2022☆13May 22, 2026Updated last month
- [CVPR 2025] FASTer: Focal Token Acquiring-and-Scaling Transformer for Long-term 3D Object Detection☆15Jun 7, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- DisTime: Distribution-based Time Representation for Video Large Language Models.☆21Jul 10, 2025Updated last year
- Seeing World Dynamics in a Nutshell☆114Mar 18, 2025Updated last year
- PyTorch implementation of FlowDiffuser: Advancing Optical Flow Estimation with Diffusion Models (CVPR-2024)☆106Jul 11, 2024Updated 2 years ago
- My implementation of a scene memory transformer module for reinforcement learning☆14Jun 19, 2019Updated 7 years ago
- The official repo for SpaceVista: All-Scale Visual Spatial Reasoning from mm to km.☆43May 26, 2026Updated last month
- [CVPR-2024] NAYER: Noisy Layer Data Generation for Efficient and Effective Data-free Knowledge Distillation☆16Oct 19, 2024Updated last year
- ☆18Jul 11, 2025Updated last year
- The source code for "LaSe-E2V: Towards Language-guided Semantic-Aware Event-to-Video Reconstruction"☆10Jul 5, 2024Updated 2 years ago
- [CVPR 2026 Highlight] MonoCoP: Unleashing the Power of Chain-of-Prediction for Monocular 3D Object Detection☆19Mar 31, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- BetterNet is a state-of-the-art deep learning model for accurate and efficient polyp segmentation in medical images. It combines Efficien…☆14May 8, 2024Updated 2 years ago
- ☆45Oct 25, 2023Updated 2 years ago
- Multi-Modal Fusion of Event and RGB for Monocular Depth Estimation Using Transformer☆19Sep 17, 2025Updated 10 months ago
- UniDoc-RL: Unified Document Understanding with Reinforcement Learning☆16May 21, 2026Updated 2 months ago
- [ICCV 2025] This is the official implementation of POMATO: Marrying Pointmap Matching with Temporal Motions for Dynamic 3D Reconstruction☆121Aug 9, 2025Updated 11 months ago
- Implementation of Learning without Prejudices: Continual Unbiased Learning via Benign and Malignant Forgetting (ICLR 2023)☆13Apr 14, 2023Updated 3 years ago
- Code for paper "PoseEmbroider:Towards a 3D, Visual, Semantic-aware Human Pose Representation" (ECCV 2024)☆18Nov 18, 2024Updated last year