Awesome Multimodal Fusion in Speech Emotion Recognition
☆17Nov 11, 2025Updated 10 months ago
Alternatives and similar repositories for awesome-multimodal-fusion-emotion-recognition
Users that are interested in awesome-multimodal-fusion-emotion-recognition are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICTC'24] - "Voice-Based Age and Gender Recognition: A Comparative Study of LSTM, RezoNet and Hybrid CNNs-BiLSTM Architecture" by Nhut Mi…☆10Jan 16, 2025Updated last year
- A multimodal SER project combining BERT and ECAPA-TDNN with cross-attention-based fusion on the IEMOCAP dataset.☆11Dec 9, 2024Updated last year
- Scientific Reports - Open access - Published: 14 February 2025☆56Oct 30, 2024Updated last year
- Cross-Speaker Encoding Network for Multi-talker Speech Recognition☆12Mar 14, 2025Updated last year
- MMER☆19Jan 8, 2026Updated 8 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code for ICASSP 2024 Paper: RECAP: Retrieval-Augmented Audio Captioning☆16Jun 23, 2024Updated 2 years ago
- Code for ICLR 2024 Paper: CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models☆23Jul 10, 2024Updated 2 years ago
- Speech emotion recognition using LSTM, SVM and MLP | 语音情感识别☆10Jul 1, 2019Updated 7 years ago
- ☆19May 19, 2024Updated 2 years ago
- ☆12Jul 29, 2022Updated 4 years ago
- ☆22Apr 22, 2024Updated 2 years ago
- Group Gated Fusion on Attention-based Bidirectional Alignment for Multimodal Emotion Recognition☆15May 10, 2022Updated 4 years ago
- ☆12Oct 12, 2024Updated last year
- A Compact and Effective Pretrained Model for Speech Emotion Recognition☆55Apr 10, 2026Updated 5 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- DWFormer: Dynamic Window Transformer for Speech Emotion Recognition(ICASSP 2023 Oral)☆68Jul 8, 2024Updated 2 years ago
- ☆15Nov 4, 2025Updated 10 months ago
- Implementation of the table detection and table structure recognition deep learning model described in the paper "ClusterTabNet: Supervis…☆13Mar 15, 2025Updated last year
- ☆13Dec 5, 2024Updated last year
- Graph to Grid: Learning Deep Representations for Multimodal Emotion Recognition☆18Apr 30, 2024Updated 2 years ago
- official repository for the paper: Multimodal emotion recognition with modality-pairwise unsupervised contrastive loss☆23Sep 6, 2022Updated 4 years ago
- ☆13Mar 23, 2026Updated 5 months ago
- 爬虫,爬取知识星球网页版☆26Jan 24, 2019Updated 7 years ago
- ☆11Nov 11, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Project focused on enhancing the quality of low-fidelity endoscopy images using Generative Adversarial Networks (GANs) implemented in PyT…☆17Jun 5, 2025Updated last year
- Pytorch implementation for the paper: Multivariate, Multi-frequency and Multimodal: Rethinking Graph Neural Networks for Emotion Recognit…☆58Dec 5, 2023Updated 2 years ago
- Extract Unique Word Lists From Wikipedia Database☆13May 27, 2020Updated 6 years ago
- The code and data for "Summary-Oriented Vision Modeling for Multimodal Abstractive Summarization"☆11May 16, 2023Updated 3 years ago
- This repo contains the code for generating the multilingual dataset introduced in the paper "MultiConAD: A Unified Multilingual Conversat…☆21Jul 9, 2025Updated last year
- ☆11Jul 16, 2024Updated 2 years ago
- Action recognition with STIP features and my own Fisher vector implementation☆14Mar 29, 2017Updated 9 years ago
- A pytorch implementation of Fine-Grained Classification via Hierarchical Bilinear Pooling with Aggregated Slack Mask (HBPASM).☆14Sep 24, 2019Updated 6 years ago
- Code for paper "Cross-Domain Slot Filling as Machine Reading Comprehension" in IJCAI 2021☆11Aug 24, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- NuNER is the family of SOTA Foundation and Zero-shot for Entity Recognition☆15Jun 11, 2024Updated 2 years ago
- ☆11Oct 24, 2022Updated 3 years ago
- Minimalist Speech-to-Text toolkit for educational purposes☆13Feb 1, 2024Updated 2 years ago
- Substitute alternative spellings of special characters (e.g. German umlauts [ae, oe, ue] and [ss]) with their correct versions (ä, ö, ü, …☆11Nov 24, 2024Updated last year
- This project is the code of BF-GCN. The paper has been accepted by IEEE Transactions on Neural Networks and Learning Systems.☆25Jul 2, 2024Updated 2 years ago
- Joint Neural Model for Entity & Relation Extraction☆16Oct 18, 2021Updated 4 years ago
- PyTorch implementation of the ICASSP-24 paper: "Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Superv…☆41Jan 6, 2024Updated 2 years ago