[ICCV 2025] The official implementation for EgoM2P: Egocentric Multimodal Multitask Pretraining.
☆38Jun 15, 2026Updated last month
Alternatives and similar repositories for EgoM2P
Users that are interested in EgoM2P are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV 2024 Oral] The official implementation of paper: COHO: Context-Sensitive City-Scale Hierarchical Urban Layout Generation☆13Aug 13, 2024Updated last year
- ☆13Mar 20, 2026Updated 4 months ago
- Official Reporsitory of "EgoMono4D: Self-Supervised Monocular 4D Scene Reconstruction for Egocentric Videos"☆48Sep 23, 2025Updated 10 months ago
- [CVPR 2025 highlight] Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision☆48Dec 2, 2025Updated 7 months ago
- Skip Mamba Diffusion for Monocular 3D Semantic Scene Completion☆12Jan 14, 2026Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Action2Sound: Ambient-Aware Generation of Action Sounds from Egocentric Videos☆26Oct 1, 2024Updated last year
- [ICCV 2025 Oral] MVTracker: Multi-view 3D Point Tracking☆512Nov 3, 2025Updated 8 months ago
- 4Deform: Neural Surface Deformation for Robust Shape Interpolation☆26Dec 4, 2025Updated 7 months ago
- [CVPR 2025] EgoLife: Towards Egocentric Life Assistant☆452Mar 19, 2025Updated last year
- Style-NeRF2NeRF implementation.☆14Dec 26, 2024Updated last year
- HD-EPIC Python script to download the entire datasets or parts of it☆24Oct 7, 2025Updated 9 months ago
- [ICCV 2025 Highlight] Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction☆437Jun 6, 2025Updated last year
- [ICCV2025] LONG3R: Long Sequence Streaming 3D Reconstruction☆44Jul 25, 2025Updated last year
- Feed-Forward SceneDINO for Unsupervised Semantic Scene Completion (ICCV 2025)☆89Sep 18, 2025Updated 10 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- The repository provides code for EgoMAN model and dataset creation scripts.☆32Dec 31, 2025Updated 6 months ago
- Official repository from the paper "Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind"☆17Mar 18, 2025Updated last year
- [CVPR 2025] GO-N3RDet: Geometry Optimized NeRF-enhanced 3D Object Detector☆16Mar 19, 2025Updated last year
- [ICCV 2025] "Fine-grained Spatiotemporal Grounding on Egocentric Videos"☆27Jul 3, 2026Updated 3 weeks ago
- [ICCV 2025] Official pytorch implementation of "SteerX: Creating Any Camera-Free 3D and 4D Scenes with Geometric Steering"☆52Mar 20, 2025Updated last year
- ☆71Jun 9, 2026Updated last month
- [CVPR'25] 🌟🌟 EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering☆52Jun 19, 2025Updated last year
- Code and data for UniEgoMotion (ICCV 2025)☆63Apr 18, 2026Updated 3 months ago
- [ICCV'25] ScenePainter: Semantically Consistent Perpetual 3D Scene Generation with Concept Relation Alignment☆37Oct 5, 2025Updated 9 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICLR 2026] PyTorch implementation of "The Less You Depend, The More You Learn: Synthesizing Novel Views from Sparse, Unposed Images with…☆66May 13, 2026Updated 2 months ago
- [ECCV 2026 Oral] One Scene, Two Depths: Probing Geometric Ambiguity in Monocular Foundation Models (Layered 3D Spatial Understanding)☆23Jul 10, 2026Updated 2 weeks ago
- DexWild: Dexterous Human Interactions for In-the-Wild Robot Policies☆45Aug 14, 2025Updated 11 months ago
- ☆12Sep 11, 2023Updated 2 years ago
- Code for "Free360: Layered Gaussian Splatting for Unbounded 360-Degree View Synthesis from Extremely Sparse and Unposed Views", CVPR 2025☆51Jul 7, 2025Updated last year
- CVPR '26 Highlight☆24May 6, 2026Updated 2 months ago
- MAPLE infuses dexterous manipulation priors from egocentric videos into vision encoders, making their features well-suited for downstream…☆34Dec 9, 2025Updated 7 months ago
- Official repo for Nymeria and NymeriaPlus datasets.☆235Jul 22, 2026Updated last week
- Official implementation of ICCV 2025 paper "EgoAgent: A Joint Predictive Agent Model in Egocentric Worlds".☆53Jun 30, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official Implementation of paper "St4RTrack: Simultaneous 4D Reconstruction and Tracking in the World"☆143Sep 18, 2025Updated 10 months ago
- Awesome Quantization Paper lists with Codes☆10Feb 24, 2021Updated 5 years ago
- [ICML 2026] 🎨 Occluded 3D Scene Reconstruction from a Single Image.☆94Jun 9, 2026Updated last month
- [WACV 2025] Official code of "SEED4D: A Synthetic Ego-Exo Dynamic 4D Data Generator, Driving Dataset and Benchmark"☆24Sep 3, 2025Updated 10 months ago
- (ICCV 2025) MonoFusion☆71Mar 22, 2026Updated 4 months ago
- [ICLR 2026 Oral (top 1.2%)] Official implementation of DepthLM☆363Jun 1, 2026Updated last month
- Code implementation of the paper 'FIction: 4D Future Interaction Prediction from Video'☆21Mar 19, 2025Updated last year