A general large multimodal model for 4D scene understanding
☆16Jul 31, 2025Updated last year
Alternatives and similar repositories for LLaVA-4D
Users that are interested in LLaVA-4D are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICRA 2024]This is the official repo of paper "HR-APR: APR-agnostic Framework with Uncertainty Estimation and Hierarchical Refinement for…☆11Feb 10, 2025Updated last year
- On solutions to the problem of Event Collapse in Motion Compensation frameworks☆15Jan 21, 2023Updated 3 years ago
- [CVPR'26] UniGame code implementation☆20Apr 21, 2026Updated 4 months ago
- DepthART official implementation☆39Oct 28, 2024Updated last year
- smplify code for point cloud based HMR☆10Jan 11, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- PyTorch implementation of E-Motion: Future Motion Simulation via Event Sequence Diffusion☆17Jul 5, 2025Updated last year
- ☆27Nov 15, 2023Updated 2 years ago
- Code release for paper EventCLIP: Adapting CLIP for Event-based Object Recognition☆25Nov 9, 2023Updated 2 years ago
- The official implementation of "Motion-prior Contrast Maximization for Dense Continuous-Time Motion Estimation" (ECCV 2024)☆27Mar 12, 2025Updated last year
- Masked Event Modeling: Self-Supervised Pretraining for Event Cameras (WACV '24)☆21Mar 19, 2024Updated 2 years ago
- Taming Contrast Maximization for Learning Sequential Event-based Optical Flow Estimation☆21Oct 2, 2023Updated 2 years ago
- PyTorch implementation of Hierarchical Neural Memory Network☆49Feb 28, 2024Updated 2 years ago
- ☆12Jul 7, 2023Updated 3 years ago
- This is a project on visual spatial reasoning tasks-SIBench☆26Jan 12, 2026Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆11Nov 27, 2025Updated 9 months ago
- A PyTorch implementation of Minimal-IK to restore SMPL parameters from skeleton/keypoints + point cloud, with VPoser pose prior for bette…☆23Dec 28, 2023Updated 2 years ago
- [CVPR 2026] Few-Shot Incremental 3D Object Detection in Dynamic Indoor Environments☆18Updated this week
- The repo for code, that hasn't been published yet☆14May 14, 2025Updated last year
- [CVPR 2025] 3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer☆101May 26, 2025Updated last year
- Lightweight Event-based Optical Flow Estimation via Iterative Deblurring☆35Mar 21, 2025Updated last year
- Source code for AEDNet paper from ACMMM 2022☆13May 22, 2026Updated 3 months ago
- DisTime: Distribution-based Time Representation for Video Large Language Models.☆22Jul 10, 2025Updated last year
- [CVPR 2025] FASTer: Focal Token Acquiring-and-Scaling Transformer for Long-term 3D Object Detection☆15Jun 7, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR 2025] The code for paper ''Video-3D LLM: Learning Position-Aware Video Representation for 3D Scene Understanding''.☆221Jun 4, 2025Updated last year
- Seeing World Dynamics in a Nutshell☆114Mar 18, 2025Updated last year
- PyTorch implementation of FlowDiffuser: Advancing Optical Flow Estimation with Diffusion Models (CVPR-2024)☆106Jul 11, 2024Updated 2 years ago
- My implementation of a scene memory transformer module for reinforcement learning☆14Jun 19, 2019Updated 7 years ago
- The official repo for SpaceVista: All-Scale Visual Spatial Reasoning from mm to km.☆44May 26, 2026Updated 3 months ago
- [CVPR-2024] NAYER: Noisy Layer Data Generation for Efficient and Effective Data-free Knowledge Distillation☆16Oct 19, 2024Updated last year
- The source code for "LaSe-E2V: Towards Language-guided Semantic-Aware Event-to-Video Reconstruction"☆10Jul 5, 2024Updated 2 years ago
- ☆46Oct 25, 2023Updated 2 years ago
- BetterNet is a state-of-the-art deep learning model for accurate and efficient polyp segmentation in medical images. It combines Efficien…☆14May 8, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR 2026 Highlight] MonoCoP: Unleashing the Power of Chain-of-Prediction for Monocular 3D Object Detection☆23Aug 3, 2026Updated 3 weeks ago
- [ICCV 2025] This is the official implementation of POMATO: Marrying Pointmap Matching with Temporal Motions for Dynamic 3D Reconstruction☆124Aug 9, 2025Updated last year
- Learning Dense and Continuous Optical Flow from an Event Camera (TIP 2022)☆62Jan 13, 2023Updated 3 years ago
- Implementation of Learning without Prejudices: Continual Unbiased Learning via Benign and Malignant Forgetting (ICLR 2023)☆13Apr 14, 2023Updated 3 years ago
- Code for paper "PoseEmbroider:Towards a 3D, Visual, Semantic-aware Human Pose Representation" (ECCV 2024)☆18Nov 18, 2024Updated last year
- Responsible Robotic Manipulation☆17Aug 31, 2025Updated 11 months ago
- ☆16Jan 27, 2026Updated 7 months ago