Code release for the paper "Egocentric Video Task Translation" (CVPR 2023 Highlight)
☆34Jun 12, 2023Updated 3 years ago
Alternatives and similar repositories for EgoT2
Users that are interested in EgoT2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- EgoTV Egocentric Task Verification from Natural Language Task Descriptions☆27Jan 9, 2024Updated 2 years ago
- ☆12Jul 22, 2025Updated last year
- Code and data release for the paper "Seeing the Arrow of Time in Large Multimodal Models"☆16Oct 2, 2025Updated 10 months ago
- Official implementation of "A Backpack Full of Skills: Egocentric Video Understanding with Diverse Task Perspectives", accepted at CVPR 2…☆24Jun 13, 2024Updated 2 years ago
- [CVPR 2023] Egocentric Audio-Visual Object Localization☆27Jan 6, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Code and data release for the paper "Learning Fine-grained View-Invariant Representations from Unpaired Ego-Exo Videos via Temporal Align…☆19Apr 5, 2024Updated 2 years ago
- Code for "Distributed, Egocentric Representations of Graphs for Detecting Critical Structures" (ICML 2019)☆20Aug 24, 2021Updated 4 years ago
- [ECCV2024] The official implementation of "Listen to Look into the Future: Audio-Visual Egocentric Gaze Anticipation".☆17Feb 24, 2025Updated last year
- Tracking Multiple Deformable Objects in Egocentric Videos (CVPR 2023)☆13Apr 10, 2023Updated 3 years ago
- Domain Adaptation and Adapters☆16Feb 28, 2023Updated 3 years ago
- Official implementation for CVPR 2025 paper "AMO Sampler: Enhancing Text Rendering with Overshooting"☆30May 3, 2025Updated last year
- Official Code for paper "Active Video Perception: Iterative Evidence Seeking for Agentic Long Video Understanding""☆18Jun 2, 2026Updated 2 months ago
- The champion solution for Ego4D Natural Language Queries Challenge in CVPR 2023☆18Jan 23, 2024Updated 2 years ago
- [CVPR 2024] "Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition"☆12Feb 27, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code implementation for paper titled "HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision"☆30Apr 16, 2024Updated 2 years ago
- ☆23Mar 20, 2024Updated 2 years ago
- ☆12Aug 7, 2024Updated 2 years ago
- Official PyTorch Implementation of the Longhorn Deep State Space Model☆56Dec 4, 2024Updated last year
- [NeurIPS 2022] Egocentric Video-Language Pretraining☆261May 9, 2024Updated 2 years ago
- ☆17Dec 22, 2025Updated 7 months ago
- Low-Computation Egocentric Barcode Detector for the Blind☆10Jun 9, 2017Updated 9 years ago
- [CVPR 2022] Egocentric Action Target Prediction in 3D☆32Dec 2, 2025Updated 8 months ago
- generative models on toys☆12Sep 10, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- CatNet: Class Incremental 3D ConvNets for Lifelong Egocentric Gesture Recognition☆12Apr 21, 2020Updated 6 years ago
- Implements the loss used in A. Furnari, S. Battiato, G. M. Farinella (2018). Leveraging Uncertainty to Rethink Loss Functions and Evaluat…☆12May 22, 2019Updated 7 years ago
- Disentangled Pre-training for Human-Object Interaction Detection☆28Sep 17, 2025Updated 10 months ago
- [ECCV 2024] VISAGE: Video Instance Segmentation with Appearance-Guided Enhancement☆39Jul 29, 2024Updated 2 years ago
- ☆27Aug 17, 2023Updated 2 years ago
- ☆11Jul 14, 2023Updated 3 years ago
- Strefer: Empowering Video LLMs with Space-Time Referring and Reasoning via Synthetic Instruction Data☆19Jun 2, 2026Updated 2 months ago
- Towards Efficient Audio-Visual Learners via Empowering Pre-trained Vision Transformers with Cross-Modal Adaptation☆15Apr 13, 2024Updated 2 years ago
- [ICCV 2023] With a Little Help from your own Past: Prototypical Memory Networks for Image Captioning.☆19Jun 7, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆12Feb 16, 2026Updated 5 months ago
- Official PyTorch implementation of AdaFlow☆66Nov 8, 2024Updated last year
- [NeurIPS‘24] Multi-Object 3D Grounding with Dynamic Modules and Language Informed Spatial Attention☆28Jun 15, 2025Updated last year
- Graph learning framework for long-term video understanding☆72Jul 13, 2026Updated last month
- A repo for processing the raw hand object detections to produce releasable pickles + library for using these☆40Oct 26, 2024Updated last year
- The official repo for "Stepping Stones: A Progressive Training Strategy for Audio-Visual Semantic Segmentation", ECCV 2024☆18Oct 11, 2024Updated last year
- [CVPR 2023] HierVL Learning Hierarchical Video-Language Embeddings☆46Aug 14, 2023Updated 3 years ago