The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS)
☆18Dec 8, 2022Updated 3 years ago
Alternatives and similar repositories for speech_emotion
Users that are interested in speech_emotion are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Depression-Detection represents a machine learning algorithm to classify audio using acoustic features in human speech, thus detecting de…☆14Jul 10, 2020Updated 6 years ago
- Source code for paper "Breaking Security-Critical Voice Authentication".☆13Jul 10, 2023Updated 3 years ago
- Thai Grapheme to Phoneme (G2P) Wiktionary Corpus☆13Jul 25, 2022Updated 3 years ago
- brainless concatenative text to speech☆16May 11, 2021Updated 5 years ago
- Lung Cancer Prediction using Machine Learning Algorithms☆20Feb 18, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Respiratory Disorder Classification Based on Lung Auscultation sounds☆13Oct 22, 2024Updated last year
- Human age estimation using deep neural networks (Keras)☆14Aug 10, 2023Updated 2 years ago
- An implementation of Speech Emotion Recognition, based on HuBERT model, training with PyTorch and HuggingFace framework, and fine-tuning …☆34May 18, 2022Updated 4 years ago
- Python interface to Optotune focus-tunable lenses☆15Feb 4, 2020Updated 6 years ago
- Open-weights voice acting pipeline combining zero-shot voice cloning with natural-language direction. Provide a reference voice (or gener…☆16May 25, 2026Updated last month
- Multi-modal Human Emotion Recognition of speech clips (audio + video) contained in RAVDESS dataset using a two stream architecture☆32Mar 2, 2023Updated 3 years ago
- A fully and partially fake speech dataset for evaluation☆15Nov 11, 2025Updated 8 months ago
- Vocal Tract Modelling by Murphy, Shelley and Ternström☆17Nov 13, 2022Updated 3 years ago
- Code for "Self-Lifting: A Novel Framework For Unsupervised Voice-Face Association Learning,ICMR,2022"☆15Oct 25, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Eye diseases classification with CNN using Pytorch 👁️☆20Feb 3, 2023Updated 3 years ago
- ☆16Apr 27, 2025Updated last year
- Official implementation of SBNet as described in "Single-branch Network for Multimodal Training".☆13Aug 28, 2023Updated 2 years ago
- Ultralytics Assets Repository☆18Aug 3, 2024Updated last year
- A powerful ComfyUI custom node that brings Google's Gemini TTS capabilities directly to your workflow. Generate high-quality speech with …☆22May 23, 2025Updated last year
- Pair Trading Analysis & Exercises Toolkit [Jupyter Notebook]☆13Nov 3, 2023Updated 2 years ago
- My implement of InstantBooth☆14Sep 11, 2023Updated 2 years ago
- ☆12Nov 1, 2023Updated 2 years ago
- A Python project for generating concatenative synthesis driven representations of audio files based on audio database analysis.☆18Feb 13, 2018Updated 8 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A rough and ready Python utility which splits audio files based on silence and desired min/max chunk duration.☆16Jun 22, 2022Updated 4 years ago
- Speaker Verification using Pytorch☆13May 23, 2024Updated 2 years ago
- Pytorch implemenation of the model proposed in the paper: Double Multi-Head Attention for Speaker Verification☆19Jul 25, 2024Updated last year
- Simple tools for ComfyUI☆18Jun 27, 2026Updated 3 weeks ago
- ☆43May 4, 2024Updated 2 years ago
- Modified LLaVA framework for MOSS2, and makes MOSS2 a multimodal model.☆13Sep 19, 2024Updated last year
- Detect emotion from audio signals of IEMOCAP dataset using multi-modal approach. Utilized acoustic features, mel-spectrogram and text as …☆41Mar 7, 2024Updated 2 years ago
- INA's library with pretrained models for gender and age prediction from faces.☆24Oct 7, 2024Updated last year
- Adafruit CircuitPython module for the MPR121 capacitive touch breakout board.☆19Apr 23, 2026Updated 2 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ChatTube: A Retrieval QA System to Youtube Videos☆10Jun 6, 2023Updated 3 years ago
- Pytorch Implementation of the Explainable Conditional Adversarial Autoencoder using Saliency Maps and SHAP (J. of Imaging - MDPI)☆12Mar 5, 2025Updated last year
- Download twitch vods, clips, and render videos with chat.☆26Jul 5, 2026Updated 2 weeks ago
- Multi-Stage Face-Voice Association Learning with Keynote Speaker Diarization (ACM MM 2024)☆22Jul 25, 2024Updated last year
- This UE4 project contains the Telekinesis Mechanic for Control☆11Jul 26, 2020Updated 5 years ago
- Implementing isometric 3D effect in a pure 2D environment.☆14Apr 21, 2021Updated 5 years ago
- CVPR2025-Multi-party Collaborative Attention Control for Image Customization☆17May 14, 2025Updated last year