ABAW3 (CVPRW): A Joint Cross-Attention Model for Audio-Visual Fusion in Dimensional Emotion Recognition
☆50Jan 15, 2024Updated 2 years ago
Alternatives and similar repositories for JointCrossAttentional-AV-Fusion
Users that are interested in JointCrossAttentional-AV-Fusion are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- FG2021: Cross Attentional AV Fusion for Dimensional Emotion Recognition☆34Nov 29, 2024Updated last year
- ABAW6 (CVPR-W) We achieved second place in the valence arousal challenge of ABAW6☆32May 21, 2024Updated 2 years ago
- ☆27Oct 7, 2021Updated 4 years ago
- ICASSP 2023: "Recursive Joint Attention for Audio-Visual Fusion in Regression Based Emotion Recognition"☆14Nov 29, 2024Updated last year
- ☆14Oct 14, 2019Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Pose-disentangled Contrastive Learning☆14Jan 27, 2024Updated 2 years ago
- Two-stage Temporal Modelling Framework for Video-based Depression Recognition using Graph Representation☆32Dec 9, 2024Updated last year
- We achieved the 2nd and 3rd places in ABAW3 and ABAW5, respectively.☆33Mar 7, 2024Updated 2 years ago
- [TCSVT'22] Official Implementation of STI-VQA☆12Oct 18, 2023Updated 2 years ago
- This repository provides implementation for the paper "Self-attention fusion for audiovisual emotion recognition with incomplete data".☆165Sep 16, 2024Updated last year
- ☆18Apr 10, 2023Updated 3 years ago
- ☆10Feb 24, 2022Updated 4 years ago
- End-to-end Multi-modal Video Temporal Grounding, NeurIPS 2021☆18Oct 24, 2021Updated 4 years ago
- ICASSP 2023: 'Speaker recognition with two-step multi-modal deep cleansing'☆44Oct 31, 2022Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The project aim was to fine-tune the stable diffusion model in order to generate images in the LEGO style based on the prompt.☆16Jun 7, 2023Updated 3 years ago
- [NeurIPS 2022] The official repository of Expression Learning with Identity Matching for Facial Expression Recognition☆45Nov 29, 2023Updated 2 years ago
- [MICCAI 2024] Feature Fusion Based on Mutual-Cross-Attention Mechanism for EEG Emotion Recognition☆60Dec 14, 2024Updated last year
- [CVPR 2024] This is the official implementation of "MART: Masked Affective RepresenTation Learning via Masked Temporal Distribution Disti…☆22Jun 14, 2025Updated last year
- Project to infere emotional expressions and benchmark datasets by Niklas Wagner, Felix Mätzler, Samed R. Vossberg, Helen Schneider and Sv…☆33Mar 7, 2025Updated last year
- [CVPR'25] AVF-MAE++ : Scaling Affective Video Facial Masked Autoencoders via Efficient Audio-Visual Self-Supervised Learning☆23Jun 11, 2026Updated 2 months ago
- ☆12May 26, 2022Updated 4 years ago
- ☆10Apr 12, 2023Updated 3 years ago
- Implementation of "A conformer-based classifier for variable-length utterance processing in anti-spoofing" published in Interspeech 2023.☆32Nov 7, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 2BiVQA is a no-reference deep learning based video quality assessment metric.☆34Sep 26, 2022Updated 3 years ago
- [ICASSP 2025] Official PyTorch code for training and inference pipeline for DepMamba: Progressive Fusion Mamba for Multimodal Depression …☆113Mar 11, 2025Updated last year
- Attention-Based Acoustic Feature Fusion Network for Depression Detection☆29Jun 14, 2025Updated last year
- 🔆 📝 A reading list focused on Multimodal Emotion Recognition (MER) 👂👄 👀 💬☆126Oct 6, 2020Updated 5 years ago
- [BMVC'23] Prompting Visual-Language Models for Dynamic Facial Expression Recognition☆143Nov 21, 2024Updated last year
- Research code for "Towards multi-task learning of speech and speaker recognition" at https://arxiv.org/pdf/2302.12773.pdf☆12Dec 2, 2024Updated last year
- soundnet and localize sound source☆12Dec 7, 2020Updated 5 years ago
- A survey of deep multimodal emotion recognition.☆57May 6, 2022Updated 4 years ago
- Code for paper "MIR-GAN: Refining Frame-Level Modality-Invariant Representations with Adversarial Network for Audio-Visual Speech Recogni…☆16Jun 21, 2023Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [AAAI'25] Official Implementation of 'MSAmba: Exploring Multimodal Sentiment Analysis with State Space Models'☆31Jul 21, 2025Updated last year
- ☆23Aug 11, 2020Updated 6 years ago
- A locally temporal-spatial pattern learning graph attention network (LTS-GAT) for EEG emotion recognition.☆12Jun 4, 2022Updated 4 years ago
- ☆71Jul 25, 2024Updated 2 years ago
- ☆19Nov 19, 2025Updated 9 months ago
- Deep Learning for Video Retrieval by Natural Language☆11Oct 20, 2019Updated 6 years ago
- A Multi-Task Evaluation Benchmark for Audio-Visual Representation Models (ICASSP 2024)☆58Apr 17, 2024Updated 2 years ago