Sharingan: A Transformer Architecture for Multi-Person Gaze Following
☆32Nov 11, 2024Updated last year
Alternatives and similar repositories for sharingan
Users that are interested in sharingan are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repo provides the training and testing code for our paper "A Modular Multimodal Architecture for Gaze Target Prediction: Application…☆25Oct 18, 2022Updated 3 years ago
- ☆13Apr 26, 2024Updated 2 years ago
- Toward Semantic Gaze Target Detection☆17Oct 24, 2025Updated 8 months ago
- ChildPlay: A New Benchmark for Understanding Children's Gaze Behaviour; code and checkpoints☆20Feb 13, 2025Updated last year
- Enhancing 3D Gaze Estimation in the Wild using Weak Supervision with Gaze Following Labels.☆19Jun 5, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Code for the CVPRW GAZE 2021 paper -- GOO : A Dataset for Gaze Object Prediction in Retail Environments☆51Apr 23, 2024Updated 2 years ago
- Official code of "ViTGaze: Gaze Following with Interaction Features in Vision Transformers"☆62Mar 3, 2025Updated last year
- An implementation of the paper "End-to-End Human-Gaze-Target Detection with Transformers"☆20Dec 5, 2024Updated last year
- The PyTorch implementation for "DEAL: Disentangle and Localize Concept-level Explanations for VLMs" (ECCV 2024 Strong Double Blind)☆20Mar 9, 2026Updated 4 months ago
- Gaze-LLE: Gaze Target Estimation via Large-Scale Learned Encoders (CVPR 2025, Highlight)☆850Mar 18, 2026Updated 4 months ago
- [CVPR 2026] GazeAnywhere: Gaze Target Estimation Anywhere with Concepts☆17Jun 3, 2026Updated last month
- Repository for 3DV2022 paper "Domain Adaptive 3D Pose Augmentation for In-the-wild Human Mesh Recovery"☆19Mar 22, 2023Updated 3 years ago
- Code for ACCV2018 paper 'Believe It or Not, We Know What You Are Looking at!'☆113Jul 9, 2021Updated 5 years ago
- Official repository of the "ReSTR: Convolution-Free Referring Image Segmentation Using Transformers (CVPR'22)"☆15Dec 13, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Paper reading: Jamba — Hybrid Transformer-Mamba LM (SSM → S4 → S6 → Jamba)☆15May 22, 2024Updated 2 years ago
- ☆54Jan 20, 2024Updated 2 years ago
- Multimodal Large Models Are Effective Action Anticipators (IEEE TMM)🌳☆27Aug 15, 2025Updated 11 months ago
- ☆23May 18, 2025Updated last year
- ☆14Jan 5, 2022Updated 4 years ago
- ☆11Oct 13, 2024Updated last year
- A Tiny JSON parser using Modern C++.☆13Jul 5, 2021Updated 5 years ago
- A lept HTTP server.☆13Jul 2, 2021Updated 5 years ago
- ContactGen: Contact-Guided Interactive 3D Human Generation for Partners (AAAI 2024)☆19Oct 11, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ECCV 2024] 3DGazeNet: Generalizing Gaze Estimation with Weak-Supervision from Synthetic Views☆126Jan 21, 2025Updated last year
- STOI loss functions in PyTorch (mirror of https://github.com/mpariente/pytorch_stoi)☆15Aug 6, 2020Updated 5 years ago
- What Do You See in Vehicle? Comprehensive Vision Solution for In-Vehicle Gaze Estimation☆52Aug 12, 2024Updated last year
- Reconstruction of highly undersampled radial cardiac MRI with a U-Net☆11Apr 4, 2020Updated 6 years ago
- Hypergraph Multi-Modal Learning for EEG-based Emotion Recognition in Conversation☆17Jun 15, 2026Updated last month
- Official implementation of "In-style: Bridging Text and Uncurated Videos with Style Transfer for Cross-modal Retrieval." ICCV 2023☆11Oct 5, 2023Updated 2 years ago
- Downscaling Intelligence: Exploring Perception and Reasoning Bottlenecks in Small Multimodal Models☆25Mar 21, 2026Updated 4 months ago
- This repository contains the Adverbs in Recipes (AIR) dataset and the code published at the CVPR 23 paper: "Learning Action Changes by Me…☆13May 25, 2023Updated 3 years ago
- ☆13Nov 28, 2021Updated 4 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Research code for CVPR 2023 paper: "Deformable mesh transformer for 3D human mesh recovery"☆31Jul 31, 2023Updated 2 years ago
- LAEO-Net++☆21Mar 24, 2021Updated 5 years ago
- AutoFuse: Automatic Fusion Networks for Unsupervised and Semi-supervised Medical Image Registration☆14Jan 8, 2025Updated last year
- Code for "Modeling Multimodal Social Interactions: New Challenges and Baselines with Densely Aligned Representations" (CVPR 2024 Oral)☆19Jun 23, 2024Updated 2 years ago
- [NeurIPS 2024] PediatricsGPT: Large Language Models as Chinese Medical Assistants for Pediatric Applications☆21Nov 4, 2024Updated last year
- [TCSVT23] Official code for "SPT: Spatial Pyramid Transformer for Image Captioning".☆10Aug 14, 2024Updated last year
- [CVPR'24] Official implementation of our paper "Self-Supervised Facial Representation Learning with Facial Region Awareness"☆15Mar 8, 2024Updated 2 years ago