☆16Oct 7, 2022Updated 3 years ago
Alternatives and similar repositories for tmh
Users that are interested in tmh are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Source code for paper Multi-Task Learning for Depression Detection in Dialogs (SIGDial 2022)☆12Jan 18, 2025Updated last year
- ☆14Feb 9, 2023Updated 3 years ago
- Adnabod lleferydd Cymraeg i'r Gymraeg gyda HuggingFace // Speech Recognition for Welsh with HuggingFace☆13Nov 29, 2022Updated 3 years ago
- Official implementation of INTERSPEECH 2021 paper 'Emotion Recognition from Speech Using Wav2vec 2.0 Embeddings'☆140Jan 6, 2025Updated last year
- ☆12Sep 25, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- AVI-R Package (formerly DIVA IO): A robust reader for AVI video files☆13Dec 21, 2020Updated 5 years ago
- The code for Multi-Scale Receptive Field Graph Model for Emotion Recognition in Conversations☆10Jan 17, 2023Updated 3 years ago
- ☆37Sep 5, 2022Updated 4 years ago
- Baseline scripts for the Audio/Visual Emotion Challenge 2019☆81Mar 5, 2022Updated 4 years ago
- Here the code of EmoAudioNet is a deep neural network for speech classification (published in ICPR 2020)☆14Jul 13, 2020Updated 6 years ago
- ICASSP 2023: "Recursive Joint Attention for Audio-Visual Fusion in Regression Based Emotion Recognition"☆14Nov 29, 2024Updated last year
- vad☆28Apr 3, 2023Updated 3 years ago
- A simple program on how you can use tensor-board for visualization and how you can freeze your model graph and later use if for testing☆14Nov 6, 2018Updated 7 years ago
- Dynamic vision-guided speaker embedding for audio-visual speaker diarization☆12Jul 5, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repository is for the paper Incorporating External POS Tagger for Punctuation Restoration. Proc. Interspeech 2021, 1987-1991, doi: 1…☆11May 24, 2026Updated 4 months ago
- Csenet: Complex Squeeze-and-Excitation Network for Speech Depression Level Prediction (ICASSP 2022)☆13Jun 23, 2022Updated 4 years ago
- ONNX Script editor & visualiser running completely in the browser thanks to Pyodide and Netron☆20Mar 29, 2023Updated 3 years ago
- This repository shows how to implement a basic model for multimodal entailment.☆10Aug 17, 2021Updated 5 years ago
- Target Agnostic Attack on Deep Models: Exploiting Security Vulnerabilities of Transfer Learning☆10Jul 2, 2019Updated 7 years ago
- [NeurIPS 2022] "Losses Can Be Blessings: Routing Self-Supervised Speech Representations Towards Efficient Multilingual and Multitask Spee…☆17Sep 19, 2023Updated 3 years ago
- Detector for faces with masks / no masks on top of them.☆18Jun 22, 2020Updated 6 years ago
- speech recognition of digits based on single Gaussian, Gaussian Mixture, and Hidden Markov Models☆11Jun 3, 2020Updated 6 years ago
- A CNN audio classifier via spectrogram images.☆10Jul 21, 2017Updated 9 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆18Jun 26, 2024Updated 2 years ago
- [ICASSP'23] This repo contains code for the Demux & MEmo emotion recognition models (https://arxiv.org/abs/2210.15842), as well as code t…☆23Jan 18, 2024Updated 2 years ago
- Stable Diffusion UI and Waifu Diffusion UI Setup Tutorial☆13Oct 1, 2022Updated 3 years ago
- Code repo for "Multi-Task Learning for Interpretable Weakly Labelled Sound Event Detection"☆17Nov 9, 2022Updated 3 years ago
- An MCP server for Kolada.☆16Nov 30, 2025Updated 9 months ago
- A Fairseq implementation of Listen, Attend and Spell (LAS), an End-to-End ASR framework.☆11Dec 21, 2020Updated 5 years ago
- An open source NLP as a service project focused on providing state of the art systems with ease. Training and inference by simple docker …☆20Sep 17, 2024Updated 2 years ago
- Code for "Distribution-based Emotion Recognition in Conversation"☆18Feb 6, 2023Updated 3 years ago
- Reference implementation and test synthetic data for Sorted Center Time echo density measure for acoustic impulse responses☆15Mar 18, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Source code and speech samples for the DSU-AVO paper accepted to INTERSPEECH 2023☆12May 13, 2024Updated 2 years ago
- ☆13Mar 25, 2021Updated 5 years ago
- ☆24Dec 10, 2022Updated 3 years ago
- A Tensorflow implementation of Speech Emotion Recognition using Audio signals and Text Data☆12May 16, 2022Updated 4 years ago
- End-to-end Speech Emotion Recognition using BLSTMs with self-attention and Multi-domain training☆49Dec 7, 2023Updated 2 years ago
- ☆18Sep 19, 2023Updated 3 years ago
- ☆24Oct 14, 2022Updated 3 years ago