Graph learning framework for long-term video understanding
☆71Jul 13, 2026Updated last month
Alternatives and similar repositories for GraVi-T
Users that are interested in GraVi-T are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- repo for active speaker detection for media videos.☆31Nov 19, 2023Updated 2 years ago
- ☆22Jul 23, 2026Updated last month
- ☆10Oct 18, 2021Updated 4 years ago
- The repository for IEEE CVPR 2023 (A Light Weight Model for Active Speaker Detection)☆187Mar 23, 2025Updated last year
- Referring expression comprehension on ReferIt(RefClef)☆10Nov 28, 2016Updated 9 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [CVPR 2022] Joint hand motion and interaction hotspots prediction from egocentric videos☆71Jan 29, 2024Updated 2 years ago
- Code release for the paper "Egocentric Video Task Translation" (CVPR 2023 Highlight)☆34Jun 12, 2023Updated 3 years ago
- INTERSPEECH2023: Target Active Speaker Detection with Audio-visual Cues☆61May 29, 2023Updated 3 years ago
- (wip) Use LAION-AI's CLIP "conditoned prior" to generate CLIP image embeds from CLIP text embeds.☆29Jul 14, 2022Updated 4 years ago
- In this codebase we establish a benchmark for egocentric user adaptation based on Ego4d.First, we start from a population model which ha…☆15Jul 24, 2026Updated last month
- [CVPR 2023] Egocentric Audio-Visual Object Localization☆27Jan 6, 2024Updated 2 years ago
- [TPAMI 2026] Learning Long-form Movie Prior via Large Language Models☆32Updated this week
- PyTorch implementation for our paper "Improving GFlowNets for Text-to-Image Diffusion Alignment."☆31Sep 6, 2024Updated last year
- ☆22Mar 7, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICCVW 2023] Interaction-Aware Prompting for Zero-Shot Spatio-Temporal Action Detection☆21Feb 22, 2024Updated 2 years ago
- [INTERSPEECH 2022] This dataset is designed for multi-modal speaker diarization and lip-speech synchronization in the wild.☆67Jan 24, 2024Updated 2 years ago
- [CVPR 2024] Official PyTorch implementation of "ECLIPSE: Revisiting the Text-to-Image Prior for Efficient Image Generation"☆65May 1, 2024Updated 2 years ago
- Implements the loss used in A. Furnari, S. Battiato, G. M. Farinella (2018). Leveraging Uncertainty to Rethink Loss Functions and Evaluat…☆12May 22, 2019Updated 7 years ago
- Tracking Multiple Deformable Objects in Egocentric Videos (CVPR 2023)☆13Apr 10, 2023Updated 3 years ago
- [T-CSVT 2021]: DeepOIS: Gyroscope-Guided Deep Optical Image Stabilizer Compensation☆11Jul 4, 2023Updated 3 years ago
- ☆23Aug 21, 2021Updated 5 years ago
- [NeurIPS 2024] Official PyTorch implementation of "Improving Compositional Reasoning of CLIP via Synthetic Vision-Language Negatives"☆49Dec 1, 2024Updated last year
- Code for "ATTA: Anomaly-aware Test-Time Adaptation for Out-of-Distribution Detection in Segmentation" (NeurIPS 23)☆16Apr 12, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- EgoTV Egocentric Task Verification from Natural Language Task Descriptions☆27Jan 9, 2024Updated 2 years ago
- The official PyTorch implementation of the IEEE/CVF Computer Vision and Pattern Recognition (CVPR) '24 paper PREGO: online mistake detect…☆35Jun 9, 2025Updated last year
- ☆20Mar 26, 2025Updated last year
- ☆14Feb 26, 2024Updated 2 years ago
- Implementation of the model: "(MC-ViT)" from the paper: "Memory Consolidation Enables Long-Context Video Understanding"☆27Updated this week
- [IJCAI-24] Explore Internal and External Similarity for Single Image Deraining with Graph Neural Networks☆11Sep 2, 2024Updated last year
- Trans4Map: Revisiting Holistic Top-down Mapping from Egocentric Images to Allocentric Semantics with Vision Transformers☆17Oct 14, 2022Updated 3 years ago
- Monaco editor (Visual Studio Code) for Streamlit☆51Sep 26, 2023Updated 2 years ago
- ☆22Nov 24, 2022Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Code and models for the Action Recognition benchmark of Assembly101☆15Mar 26, 2023Updated 3 years ago
- Code for "Distributed, Egocentric Representations of Graphs for Detecting Critical Structures" (ICML 2019)☆20Aug 24, 2021Updated 5 years ago
- This repository is for The Power of Sound(TPoS): Audio Reactive Video Generation with Stable Diffusion (ICCV2023)☆25Dec 7, 2023Updated 2 years ago
- A robust PCA method of tumor clone and evolution inference from single-cell sequencing data.☆12May 28, 2020Updated 6 years ago
- ☆23Mar 20, 2024Updated 2 years ago
- Interface to stable-baselines3 APIs for training RL policies on gym-registered environments☆12Jan 24, 2024Updated 2 years ago
- Python package for egocentric network analysis☆14Feb 6, 2018Updated 8 years ago