Building Egocentric Procedural AI Assistant: Methods, Benchmarks, and Challenges
☆53Jul 23, 2026Updated this week
Alternatives and similar repositories for Building-Egocentric-Procedural-AI-Assistant
Users that are interested in Building-Egocentric-Procedural-AI-Assistant are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- NeurIPS 2025☆19Oct 20, 2025Updated 9 months ago
- [ICLR 2025] OccProphet: Pushing Efficiency Frontier of Camera-Only 4D Occupancy Forecasting with Observer-Forecaster-Refiner Framework☆60Mar 18, 2026Updated 4 months ago
- ☆18Apr 5, 2024Updated 2 years ago
- ☆12Jul 22, 2025Updated last year
- Vinci: A Real-time Embodied Smart Assistant based on Egocentric Vision-Language Model☆93Nov 27, 2025Updated 8 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ProactiveBench: A Comprehensive Benchmark for VideoLLM Proactive Interaction Evaluation☆20Jan 8, 2026Updated 6 months ago
- [RAL 2022] S2G2: Semi-Supervised Semantic Bird-Eye-View Grid-Map Generation Using a Monocular Camera for Autonomous Driving☆11Nov 23, 2022Updated 3 years ago
- Multimodaler Anomalie-Detektions Benchmark für simulierte Szenarien☆15Jul 16, 2024Updated 2 years ago
- Fragments-Expert is a software package for feature extraction from file fragments and classification among various file formats.☆13Jan 16, 2024Updated 2 years ago
- [NeurIPS'23] The official implementation of paper "Bitstream-corrupted Video Recovery: A Novel Benchmark Dataset and Method"☆44Jul 25, 2025Updated last year
- About [MM2024] Learning with Alignments: Tackling the Inter- and Intra-domain Shifts for Cross-multidomain Facial Expression Recognition☆16Nov 12, 2024Updated last year
- ☆11Jul 14, 2023Updated 3 years ago
- [TIV 2025] C2L-PR: Cross-modal Camera-to-LiDAR Place Recognition via Modality Alignment and Orientation Voting.☆20Mar 28, 2026Updated 4 months ago
- STI-Bench : Are MLLMs Ready for Precise Spatial-Temporal World Understanding?☆39Jan 12, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆19Apr 11, 2026Updated 3 months ago
- The official implementation of Error Detection in Egocentric Procedural Task Videos☆33Sep 20, 2025Updated 10 months ago
- [CVPR'25] 🌟🌟 EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering☆52Jun 19, 2025Updated last year
- Master AI Agent Assistant in 3 Days. A guided study plan using nanobot (~3k lines of Python) and your own AI Socratic Tutor. Learn Archit…☆15Feb 7, 2026Updated 5 months ago
- ☆13Jun 13, 2025Updated last year
- Official repository from the paper "Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind"☆17Mar 18, 2025Updated last year
- ☆19Jun 13, 2025Updated last year
- ☆40Nov 5, 2025Updated 8 months ago
- MSFSR:A Multi-Stage Face Super-Resolution with Accurate Facial Representation via Enhanced Facial Boundaries☆12Jun 15, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Fully Decoupled Neural Network Learning Using Delayed Gradients (FDG)☆21Jul 5, 2021Updated 5 years ago
- [ICME2024, Official Code] for paper "Bringing Textual Prompt to AI-Generated Image Quality Assessment"☆21Jul 9, 2024Updated 2 years ago
- Code for "Agentic Very Long Video Understanding" (EGAgent) [ACL 2026 Main]☆50Jul 1, 2026Updated 3 weeks ago
- Visual Relationship Reasoning for Grasp Planning☆19May 22, 2025Updated last year
- [CVPR2024] OHTA: One-shot Hand Avatar via Data-driven Implicit Priors☆33Jun 14, 2024Updated 2 years ago
- BEAR: a new BEnchmark on video Action Recognition☆46Apr 21, 2024Updated 2 years ago
- Code and data for UniEgoMotion (ICCV 2025)☆63Apr 18, 2026Updated 3 months ago
- Feature SIMilarity Index for Tone Mapping - A perceptual image quality assessment tool☆10Mar 6, 2018Updated 8 years ago
- Fine-tune Qwen2.5-VL-7B on custom visual QA tasks using LoRA + Accelerate, supporting single/multi-GPU training on COCO 2014 dataset.☆30Apr 28, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- PyTorch code for BMVC 2018 ``Strong Baseline for Single Image Dehazing with Deep Features and Instance Normalization''☆20Nov 5, 2018Updated 7 years ago
- Python partial re-implementation of accumuLaser in python from the KITTI360 devkits to recover label of individual pointclouds from aggre…☆32Jan 29, 2022Updated 4 years ago
- The Pytorch reproduction of WMCNN [Aerial Image Super Resolution via Wavelet Multiscale Convolutional Neural Networks]☆18Aug 19, 2020Updated 5 years ago
- Joint Learning Content and Degradation Aware Embedding for Blind Super-Resolution☆14Oct 20, 2022Updated 3 years ago
- [CVPR 2026] Implementation of HAMMER: Harnessing MLLMs via Cross-Modal Integration for Intention-Driven 3D Affordance Grounding☆20Apr 30, 2026Updated 2 months ago
- Code release for the paper "Egocentric Video Task Translation" (CVPR 2023 Highlight)☆34Jun 12, 2023Updated 3 years ago
- [CVPR 2025 Highlight] This repo is official PyTorch implementation of End-to-End HOI Reconstruction Transformer with Graph-based Encoding…☆25May 8, 2025Updated last year