Official Repository for "Learning Trimodal Relation for Audio-Visual Question Answering with Missing Modality" (ECCV 2024)
☆16Oct 29, 2024Updated last year
Alternatives and similar repositories for Missing-AVQA
Users that are interested in Missing-AVQA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Repository for "Audio-Visual Spatial Integration and Recursive Attention for Robust Sound Source Localization" (ACM MM 2023)☆18Nov 14, 2023Updated 2 years ago
- Official Repository for "Learning to Visually Localize Sound Sources from Mixtures without Prior Source Knowledge" (CVPR 2024)☆17Sep 1, 2024Updated last year
- ☆17Aug 11, 2023Updated 3 years ago
- Official Repository for "Multispectral Pedestrian Detection with Sparsely Annotated Label" (AAAI 2025)☆32Apr 28, 2025Updated last year
- [2025 CVPR] Towards Open-Vocabulary Audio-Visual Event Localization☆46Mar 7, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The repo for "On-the-fly Modulation for Balanced Multimodal Learning", T-PAMI 2024☆19Sep 29, 2024Updated last year
- This repository contains code for AAAI2025 paper "Dense Audio-Visual Event Localization under Cross-Modal Consistency and Multi-Temporal …☆26Aug 18, 2025Updated 11 months ago
- Official codebase for "Context Aware Deep Learning for Multi Modal Depression Detection" [ICASSP 2019, Oral]☆11Dec 26, 2024Updated last year
- [AAAI 2023 Oral] Official pytorch implementation of "Towards Good Practices for Missing Modality Robust Action Recognition"☆23Dec 1, 2022Updated 3 years ago
- [CVPR 2025] 🔥 Official impl. of "Audio-Visual Instance Segmentation".☆52Jun 5, 2025Updated last year
- This is the repo for "Adaptive Unimodal Regulation for Balanced Multimodal Information Acquisition", CVPR2025.☆26Dec 22, 2025Updated 7 months ago
- MUSIC-AVQA, CVPR2022 (ORAL)☆100Dec 30, 2022Updated 3 years ago
- ☆25Apr 16, 2025Updated last year
- ☆13May 21, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- An official implementation of "Incomplete Multimodality-Diffused Emotion Recognition" in PyTorch. (NeurIPS 2023)☆65Dec 5, 2023Updated 2 years ago
- The official code repository of ShaSpec model from CVPR 2023 [paper](https://arxiv.org/pdf/2307.14126) "Multi-modal Learning with Missing…☆101Apr 16, 2025Updated last year
- ☆36Jul 25, 2024Updated 2 years ago
- Official repository for "Boosting Audio Visual Question Answering via Key Semantic-Aware Cues" in ACM MM 2024.☆17Oct 25, 2024Updated last year
- [2024 ECCV] Label-anticipated Event Disentanglement for Audio-Visual Video Parsing☆14Nov 17, 2024Updated last year
- A Lightweight Multi-modality Image Segmentation Network via Domain Adaptation using Gradient Magnitude and Shape Constraint☆10Apr 3, 2023Updated 3 years ago
- ☆14Aug 13, 2020Updated 6 years ago
- ☆29Aug 2, 2023Updated 3 years ago
- [CVPR 2024 CVinW] Multi-Agent VQA: Exploring Multi-Agent Foundation Models on Zero-Shot Visual Question Answering☆22Sep 21, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆23Apr 19, 2024Updated 2 years ago
- ☆41Apr 16, 2024Updated 2 years ago
- This repository provides the codes for MMA-DFER: multimodal (audiovisual) emotion recognition method. This is an official implementation …☆57Sep 16, 2024Updated last year
- ☆12Apr 19, 2024Updated 2 years ago
- [IJCNN 2021] FedCM: A Real-time Contribution Measurement Method for Participants in Federated Learning☆11Aug 21, 2021Updated 4 years ago
- SwinTransformer for Tensorflow2☆11Jul 7, 2022Updated 4 years ago
- [MICCAI 2023] GRACE: Enhancing Federated Learning for Medical Imaging with Generalized and Personalized Gradient Correction☆17Jun 29, 2023Updated 3 years ago
- The benchmark for "Video Object Segmentation in Panoptic Wild Scenes".☆12Oct 17, 2023Updated 2 years ago
- 一款即插即用的知识蒸馏工具包☆13May 16, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code for ''A Simple Baseline for Audio-Visual Scene-Aware Dialog``☆27May 26, 2020Updated 6 years ago
- The official repository of the paper "InfoCD: A Contrastive Chamfer Distance Loss for Point Cloud Completion" published at NeurIPS 2023☆23Oct 13, 2023Updated 2 years ago
- RA-Touch: Retrieval-Augmented Touch Understanding with Enriched Visual Data (ACM MM '25)☆15Sep 12, 2025Updated 11 months ago
- ☆11Aug 20, 2025Updated 11 months ago
- [ECCVW 2022] UAD: Localization Uncertainty Estimation for Anchor-Free Object Detection☆15Aug 1, 2023Updated 3 years ago
- Code for A Dual Domain Multi-exposure Image Fusion Network Based on the Spatial-frequency Integration.☆12Jul 25, 2024Updated 2 years ago
- [TGRS 2023] Point Label Meets Remote Sensing Change Detection: A Consistency-Aligned Regional Growth Network☆15Jan 5, 2024Updated 2 years ago