For Ego4D VQ3D Task
☆22Jan 9, 2024Updated 2 years ago
Alternatives and similar repositories for EgoLoc
Users that are interested in EgoLoc are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- CVPR 2024 "Instance Tracking in 3D Scenes from Egocentric Videos"☆19Jun 27, 2024Updated 2 years ago
- An experiment with movie scenes and contrastive learning☆11Feb 1, 2025Updated last year
- Code release for "EgoVLPv2: Egocentric Video-Language Pre-training with Fusion in the Backbone" [ICCV, 2023]☆110Jul 2, 2024Updated 2 years ago
- Human-centric environment representations from egocentric video☆15Feb 5, 2026Updated 5 months ago
- Official PyTorch code of GroundVQA (CVPR'24)☆63Sep 13, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [AAAI 2024] UniAP: Towards Universal Animal Perception in Vision via Few-shot Learning☆12Dec 10, 2023Updated 2 years ago
- [CHI24] AI-Assisted In-Context Writing on OHMD During Travels☆12Dec 19, 2024Updated last year
- EventEgo3D: 3D Human Motion Capture from Egocentric Event Streams [CVPR'24]☆32Jul 23, 2025Updated last year
- Code and data release for the paper "Learning Object State Changes in Videos: An Open-World Perspective" (CVPR 2024)☆37Sep 9, 2024Updated last year
- ☆29Jun 29, 2026Updated last month
- ☆13Nov 5, 2024Updated last year
- FleVRS: Towards Flexible Visual Relationship Segmentation, NeurIPS 2024☆22Dec 9, 2024Updated last year
- A curated list of egocentric (first-person) vision and related area resources☆335Jul 6, 2026Updated 3 weeks ago
- The code for "VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by VIdeo SpatioTemporal Augmentation" [CVPR2025]☆20Feb 27, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICCV 2023] Global Adaptation meets Local Generalization: Unsupervised Domain Adaptation for 3D Human Pose Estimation☆24Aug 26, 2023Updated 2 years ago
- Official implementation of `Discovering Hidden Visual Concepts Beyond Linguistic Input in Infant Learning`, CVPR 2025☆12Aug 1, 2025Updated 11 months ago
- ☆53Oct 27, 2023Updated 2 years ago
- ☆11Jul 26, 2018Updated 8 years ago
- Official implementation of "A Backpack Full of Skills: Egocentric Video Understanding with Diverse Task Perspectives", accepted at CVPR 2…☆24Jun 13, 2024Updated 2 years ago
- A Massive Multi-Discipline Lecture Understanding Benchmark☆34Apr 20, 2026Updated 3 months ago
- ☆15Dec 6, 2023Updated 2 years ago
- 📚 A collection of resources and papers on Large Language Models in autonomous driving☆27Oct 30, 2023Updated 2 years ago
- Source code for the paper "FESTA: Flow Estimation via Spatial-Temporal Attention for Scene Point Clouds"☆18Sep 13, 2021Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The official code and data for paper "VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI"☆18Mar 25, 2025Updated last year
- ☆20Apr 14, 2023Updated 3 years ago
- ☆37Sep 16, 2024Updated last year
- UGaitNet: Multimodal gait recognition with missing input modalities☆13May 24, 2022Updated 4 years ago
- Human-centered Delivery Benchmark☆20Jul 24, 2024Updated 2 years ago
- [NeurIPS 2022] Egocentric Video-Language Pretraining☆261May 9, 2024Updated 2 years ago
- [CVPR 2023] Code for the paper "Masked Images Are Counterfactual Samples for Robust Fine-tuning"☆14Mar 24, 2023Updated 3 years ago
- ☆139May 30, 2024Updated 2 years ago
- ☆162Aug 23, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for the paper "Vamba: Understanding Hour-Long Videos with Hybrid Mamba-Transformers" [ICCV 2025]☆107Jul 28, 2025Updated last year
- ☆12Dec 15, 2023Updated 2 years ago
- [ICLR'25] Do Egocentric Video-Language Models Truly Understand Hand-Object Interactions?☆14Apr 11, 2025Updated last year
- ☆14Nov 28, 2022Updated 3 years ago
- [CVPR 2022] Sequential Voting with Relational Box Fields for Active Object Detection☆10Jun 19, 2022Updated 4 years ago
- [ECCV 2022] GEB+: A Benchmark for Generic Event Boundary Captioning, Grounding and Retrieval☆17Aug 24, 2022Updated 3 years ago
- [IROS 24] Official repository of "Mind the Error! Detection and Localization of Instruction Errors in Vision-and-Language Navigation". We…☆19Apr 1, 2026Updated 3 months ago