IEEE T-BIOM : "Audio-Visual Fusion for Emotion Recognition in the Valence-Arousal Space Using Joint Cross-Attention"
☆48Nov 29, 2024Updated last year
Alternatives and similar repositories for Joint-Cross-Attention-for-Audio-Visual-Fusion
Users that are interested in Joint-Cross-Attention-for-Audio-Visual-Fusion are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- FG2021: Cross Attentional AV Fusion for Dimensional Emotion Recognition☆34Nov 29, 2024Updated last year
- PyTorch implementation for Audio-Visual Domain Adaptation Feature Fusion for Speech Emotion Recognition☆12Mar 20, 2022Updated 4 years ago
- ICASSP 2023: "Recursive Joint Attention for Audio-Visual Fusion in Regression Based Emotion Recognition"☆14Nov 29, 2024Updated last year
- This repository provides implementation for the paper "Self-attention fusion for audiovisual emotion recognition with incomplete data".☆168Sep 16, 2024Updated 2 years ago
- ABAW6 (CVPR-W) We achieved second place in the valence arousal challenge of ABAW6☆32May 21, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- code for paper 'Spatial-Temporal Attention Network for Depression Recognition from Facial Videos'☆37Dec 27, 2024Updated last year
- "MULTIMODAL EMOTION RECOGNITION BASED ON DEEP TEMPORAL FEATURES USING CROSS-MODAL TRANSFORMER AND SELF-ATTENTION" ICASSP'23☆24Feb 26, 2023Updated 3 years ago
- Two-stage Temporal Modelling Framework for Video-based Depression Recognition using Graph Representation☆34Dec 9, 2024Updated last year
- ☆14Oct 10, 2024Updated last year
- [Information Fusion 2024] HiCMAE: Hierarchical Contrastive Masked Autoencoder for Self-Supervised Audio-Visual Emotion Recognition☆122Aug 29, 2025Updated last year
- Multimodal SER Model meant to be trained on recognising emotions from speech (text + acoustic data). Fine-tuned the DeBERTaV3 model, resp…☆11Jun 19, 2024Updated 2 years ago
- We present a study of a neural network based method for speech emotion recognition, using audio-only features. In the studied scheme, the…☆11Jul 24, 2024Updated 2 years ago
- This repository provides the ability to recoginize the emotion from video using audiovisual modalities。端到端的多模态情感识别代码☆11Mar 5, 2023Updated 3 years ago
- FRAME-LEVEL EMOTIONAL STATE ALIGNMENT METHOD FOR SPEECH EMOTION RECOGNITION☆22Dec 22, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- a deep learning method to detect 68 landmarks☆15Aug 1, 2018Updated 8 years ago
- [TCSVT'22] Official Implementation of STI-VQA☆12Oct 18, 2023Updated 2 years ago
- 多模态,语音和文本结合的情感识别,大模型finetune☆25Nov 19, 2023Updated 2 years ago
- A survey of deep multimodal emotion recognition.☆57May 6, 2022Updated 4 years ago
- The PyTorch code for paper: "CONSK-GCN: Conversational Semantic- and Knowledge-Oriented Graph Convolutional Network for Multimodal Emotio…☆13Oct 21, 2022Updated 3 years ago
- A Fully End2End Multimodal System for Fast Yet Effective Video Emotion Recognition☆39Aug 12, 2024Updated 2 years ago
- A wrapper for Audeering's wav2vec-based dimensional speech emotion recognition☆22Aug 9, 2023Updated 3 years ago
- ☆22Apr 22, 2024Updated 2 years ago
- ☆39Jun 28, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repository contains the code for the paper: "DeToxy: A Large-Scale Multimodal Dataset for Toxicity Classification in Spoken Utteranc…☆21Oct 13, 2022Updated 3 years ago
- AD-TUNING: An Adaptive CHILD-TUNING Approach to Efficient Hyperparameter Optimization of Child Networks for Speech Processing Tasks in th…☆11Feb 23, 2024Updated 2 years ago
- Code for MMLatch: Bottom-up Top-down Fusion for Multimodal Sentiment Analysis https://arxiv.org/abs/2201.09828 (to be presented in ICASSP…☆34Mar 7, 2023Updated 3 years ago
- official repository for the paper: Multimodal emotion recognition with modality-pairwise unsupervised contrastive loss☆23Sep 6, 2022Updated 4 years ago
- [IEEE ICPRS 2024 Oral] TensorFlow code implementation of "MultiMAE-DER: Multimodal Masked Autoencoder for Dynamic Emotion Recognition"☆19Mar 13, 2026Updated 6 months ago
- Multimodal emotion recognition system of attention based vision network + audio network☆14Jul 21, 2020Updated 6 years ago
- Pytorch implementation for the paper: Multivariate, Multi-frequency and Multimodal: Rethinking Graph Neural Networks for Emotion Recognit…☆58Dec 5, 2023Updated 2 years ago
- Multi-modal Human Emotion Recognition of speech clips (audio + video) contained in RAVDESS dataset using a two stream architecture☆32Mar 2, 2023Updated 3 years ago
- This repository provides the codes for MMA-DFER: multimodal (audiovisual) emotion recognition method. This is an official implementation …☆57Sep 16, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆20Aug 22, 2024Updated 2 years ago
- Trustworthy Speech Emotion Recognition☆13May 22, 2023Updated 3 years ago
- ☆24Apr 16, 2025Updated last year
- ☆85Dec 4, 2024Updated last year
- Offical implementation of paper "MSAF: Multimodal Split Attention Fusion"☆80Jun 16, 2021Updated 5 years ago
- DWFormer: Dynamic Window Transformer for Speech Emotion Recognition(ICASSP 2023 Oral)☆68Jul 8, 2024Updated 2 years ago
- Auto Generate Subtitle File For Any Type Of Audio and Video. Using Python and Google Speech-to-Text API.☆13May 15, 2020Updated 6 years ago