[ECCV 2022] AssistQ: Affordance-centric Question-driven Task Completion for Egocentric Assistant
☆23Jan 30, 2026Updated 7 months ago
Alternatives and similar repositories for Q2A
Users that are interested in Q2A are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [AAAI2023] Symbolic Replay: Scene Graph as Prompt for Continual Learning on VQA Task (Oral)☆43Mar 23, 2024Updated 2 years ago
- Edit and Generate Anything in 3D world!☆13Apr 15, 2023Updated 3 years ago
- Code for MANO-GCN —— "Capturing Implicit Spatial Cues for Monocular 3D Hand Reconstruction" (ICME2021 Oral)☆13Jun 24, 2021Updated 5 years ago
- This is the project page for the HOSNeRF☆16Dec 11, 2023Updated 2 years ago
- Code accompanying EGO-TOPO: Environment Affordances from Egocentric Video (CVPR 2020)☆31Aug 3, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- team Doggeee's solution to Ego4D LTA challenge@CVPRW23'☆14Nov 4, 2023Updated 2 years ago
- NaQ: Leveraging Narrations as Queries to Supervise Episodic Memory. CVPR 2023.☆17Jan 26, 2024Updated 2 years ago
- [NeurIPS 2022] Egocentric Video-Language Pretraining☆262May 9, 2024Updated 2 years ago
- ☆75May 10, 2024Updated 2 years ago
- [ICCV 2023] Label-Efficient Online Continual Object Detection in Streaming Video☆23Jan 8, 2024Updated 2 years ago
- [ICCV 2025] Balanced Image Stylization with Style Matching Score☆70Mar 9, 2026Updated 6 months ago
- ☆140Mar 16, 2023Updated 3 years ago
- (ECCV 2024) Empowering Multimodal Large Language Model as a Powerful Data Generator☆116Mar 21, 2025Updated last year
- In this codebase we establish a benchmark for egocentric user adaptation based on Ego4d.First, we start from a population model which ha…☆15Jul 24, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML 2026]A Large-scale Dataset for training and evaluating model's ability on Dense Text Image Generation☆94Sep 27, 2025Updated 11 months ago
- ☆14Jun 15, 2022Updated 4 years ago
- Code for the paper "GenHowTo: Learning to Generate Actions and State Transformations from Instructional Videos" published at CVPR 2024☆54Mar 3, 2024Updated 2 years ago
- [Findings of EMNLP 2022] AssistSR: Task-oriented Video Segment Retrieval for Personal AI Assistant☆24Sep 11, 2023Updated 3 years ago
- ☆62Apr 28, 2025Updated last year
- Code for NeurIPS 2022 Datasets and Benchmarks paper - EgoTaskQA: Understanding Human Tasks in Egocentric Videos.☆47Apr 17, 2023Updated 3 years ago
- ☆86Mar 4, 2024Updated 2 years ago
- The official implementation of paper: Estimating Egocentric 3D Human Pose in Global Space.☆13Sep 23, 2023Updated 3 years ago
- T2VScore: Towards A Better Metric for Text-to-Video Generation☆81Apr 10, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- (wip) Use LAION-AI's CLIP "conditoned prior" to generate CLIP image embeds from CLIP text embeds.☆29Jul 14, 2022Updated 4 years ago
- ☆13Jul 6, 2022Updated 4 years ago
- PIC API☆25Sep 18, 2019Updated 7 years ago
- Low-Computation Egocentric Barcode Detector for the Blind☆10Jun 9, 2017Updated 9 years ago
- DoraCycle: Domain-Oriented Adaptation of Unified Generative Model in Multimodal Cycles☆31Mar 8, 2026Updated 6 months ago
- [ICCV 2023] Understanding 3D Object Interaction from a Single Image☆47Feb 29, 2024Updated 2 years ago
- CatNet: Class Incremental 3D ConvNets for Lifelong Egocentric Gesture Recognition☆12Apr 21, 2020Updated 6 years ago
- ☆66Jun 16, 2023Updated 3 years ago
- ☆12Apr 6, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆23May 5, 2026Updated 4 months ago
- ManipulaTHOR, a framework that facilitates visual manipulation of objects using a robotic arm☆99Feb 7, 2023Updated 3 years ago
- Improvements made to pietrolechthaler's and his group project titled: "UR5 Pick and Place Simulation in Ros/Gazebo", available in the nex…☆11May 10, 2023Updated 3 years ago
- FQGAN: Factorized Visual Tokenization and Generation☆59Mar 29, 2025Updated last year
- [ICRA2023] Grounding Language with Visual Affordances over Unstructured Data☆49Oct 29, 2023Updated 2 years ago
- Trans4Map: Revisiting Holistic Top-down Mapping from Egocentric Images to Allocentric Semantics with Vision Transformers☆17Oct 14, 2022Updated 3 years ago
- Code and models for the Action Recognition benchmark of Assembly101☆16Mar 26, 2023Updated 3 years ago