Code for "Modeling Multimodal Social Interactions: New Challenges and Baselines with Densely Aligned Representations" (CVPR 2024 Oral)
☆20Jun 23, 2024Updated 2 years ago
Alternatives and similar repositories for MMSI
Users that are interested in MMSI are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV2024] Nonverbal Interaction Detection☆32Oct 30, 2024Updated last year
- [ACL2023] Source codes for the paper "Werewolf Among Us: Multimodal Resources for Modeling Persuasion Behaviors in Social Deduction Games…☆16Feb 22, 2025Updated last year
- Robustly Converting Camera View from Normal View to Top View for Autonomous Vehicle System on Robotics Operating System (ROS)☆24Jan 29, 2020Updated 6 years ago
- RIT-18: A Novel Dataset for Compositional Group Activity Understanding☆11Jun 15, 2020Updated 6 years ago
- [CVPR 2026] GazeAnywhere: Gaze Target Estimation Anywhere with Concepts☆23Aug 11, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ECCV2024] The official implementation of "Listen to Look into the Future: Audio-Visual Egocentric Gaze Anticipation".☆18Feb 24, 2025Updated last year
- ☆24Jul 23, 2026Updated 2 months ago
- ☆23Mar 7, 2025Updated last year
- Official Repository for "Audio-Visual Spatial Integration and Recursive Attention for Robust Sound Source Localization" (ACM MM 2023)☆18Nov 14, 2023Updated 2 years ago
- Official PyTorch implementation of "Towards More Practical Group Activity Detection: A New Benchmark and Model"☆18Oct 29, 2024Updated last year
- ☆24Jul 1, 2025Updated last year
- Official source code for the paper "Tailored Design of Audio-Visual Speech Recognition Models using Branchformers"☆15Feb 24, 2025Updated last year
- 【CVPR2023】GFIE: A Dataset and Baseline for Gaze-Following from 2D to 3D in Indoor Environments☆34Oct 16, 2023Updated 2 years ago
- The benchmark for "Video Object Segmentation in Panoptic Wild Scenes".☆13Oct 17, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Pytorch implementation of "Towards Practical and Efficient Image-to-Speech Captioning with Vision-Language Pre-training and Multi-modal T…☆12Apr 29, 2026Updated 5 months ago
- ☆35Aug 26, 2024Updated 2 years ago
- [CVPR 2022] Sequential Voting with Relational Box Fields for Active Object Detection☆10Jun 19, 2022Updated 4 years ago
- The official implementation of the paper "Affective Faces for Goal-Driven Dyadic Communication."☆15Jan 27, 2023Updated 3 years ago
- STOI loss functions in PyTorch (mirror of https://github.com/mpariente/pytorch_stoi)☆15Aug 6, 2020Updated 6 years ago
- (NeXD @ CVPR 2025) Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Models☆34Sep 30, 2025Updated last year
- Reconstruction of highly undersampled radial cardiac MRI with a U-Net☆11Apr 4, 2020Updated 6 years ago
- Mental state inference from observable behavior☆15Dec 3, 2021Updated 4 years ago
- Official Code Repository for the paper "Generating Realistic Images from In-the-wild Sounds", ICCV 2023☆12Aug 24, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆10Aug 22, 2023Updated 3 years ago
- [ECCV2024, Oral, Best Paper Finalist] This is the official implementation of the paper "LEGO: Learning EGOcentric Action Frame Generation…☆41Feb 24, 2025Updated last year
- ☆13Jul 20, 2022Updated 4 years ago
- Official repo of the paper "Object-aware Gaze Target Detection" (ICCV 2023)☆46Jul 23, 2026Updated 2 months ago
- Code for paper "RapVerse: Coherent Vocals and Whole-Body Motions Generations from Text"☆18May 30, 2024Updated 2 years ago
- Efficient Training for Multilingual Visual Speech Recognition: Pre-training with Discretized Visual Speech Representation (ACM MM 2024)☆20Mar 17, 2025Updated last year
- LAEO-Net++☆21Mar 24, 2021Updated 5 years ago
- Task-Focused Few-Shot Object Detection Benchmark☆14Jun 24, 2025Updated last year
- This repo contains script to download MUSIC dataset from youtube☆13Jan 19, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- AutoFuse: Automatic Fusion Networks for Unsupervised and Semi-supervised Medical Image Registration☆15Jan 8, 2025Updated last year
- [AAAI 2024]Weakly Supervised Multimodal Affordance Grounding for Egocentric Images☆13Nov 10, 2024Updated last year
- [NeurIPS 2024] PediatricsGPT: Large Language Models as Chinese Medical Assistants for Pediatric Applications☆23Nov 4, 2024Updated last year
- [CVPR'24] Official implementation of our paper "Self-Supervised Facial Representation Learning with Facial Region Awareness"☆15Mar 8, 2024Updated 2 years ago
- a math-formula image recognition project which placed at the first place in a competition hosted by NAVER CONNECT boostcamp AI Tech☆10Dec 16, 2023Updated 2 years ago
- This is the official implementation of the ICML 2023 paper "Fair yet Asymptotically Equal Collaborative Learning"☆10May 29, 2023Updated 3 years ago
- Generate custom text files for dataloader within UDA methods☆14May 24, 2023Updated 3 years ago