The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS)
☆19Dec 8, 2022Updated 3 years ago
Alternatives and similar repositories for speech_emotion
Users that are interested in speech_emotion are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Here are some programs made with Python and JavaScript (p5.js) related to artificial intelligence.☆10Jun 19, 2020Updated 6 years ago
- Thai Grapheme to Phoneme (G2P) Wiktionary Corpus☆13Jul 25, 2022Updated 4 years ago
- Tic Tac Toe game with socket programming and pygame☆10Jan 6, 2024Updated 2 years ago
- Tools for the automatic detection of speech-related inhalation events and characterisation of the speech respiratory cycle.☆11Feb 17, 2024Updated 2 years ago
- Respiratory Disorder Classification Based on Lung Auscultation sounds☆13Oct 22, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Human age estimation using deep neural networks (Keras)☆14Aug 10, 2023Updated 3 years ago
- A versatile, easily configurable vocoder software in MATLAB, for research purposes☆15Apr 9, 2021Updated 5 years ago
- A fully and partially fake speech dataset for evaluation☆15Nov 11, 2025Updated 10 months ago
- Vocal Tract Modelling by Murphy, Shelley and Ternström☆17Nov 13, 2022Updated 3 years ago
- Open-weights voice acting pipeline combining zero-shot voice cloning with natural-language direction. Provide a reference voice (or gener…☆18May 25, 2026Updated 3 months ago
- Code for "Self-Lifting: A Novel Framework For Unsupervised Voice-Face Association Learning,ICMR,2022"☆15Oct 25, 2024Updated last year
- ☆16Apr 27, 2025Updated last year
- Github repo for MARVEL: Multidimensional Abstraction and Reasoning through Visual Evaluation and Learning☆18Jun 12, 2024Updated 2 years ago
- A powerful ComfyUI custom node that brings Google's Gemini TTS capabilities directly to your workflow. Generate high-quality speech with …☆23May 23, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Pair Trading Analysis & Exercises Toolkit [Jupyter Notebook]☆13Nov 3, 2023Updated 2 years ago
- My implement of InstantBooth☆14Sep 11, 2023Updated 3 years ago
- A three-dimensional vocal tract acoustic model using the finite-difference time-domain (FDTD) numerical scheme.☆20Sep 25, 2022Updated 3 years ago
- ☆12Nov 1, 2023Updated 2 years ago
- A rough and ready Python utility which splits audio files based on silence and desired min/max chunk duration.☆16Jun 22, 2022Updated 4 years ago
- Biostatistics Guide☆15Jul 20, 2021Updated 5 years ago
- Python interface to Optotune focus-tunable lenses☆15Feb 4, 2020Updated 6 years ago
- A PyTorch implementation of MIT CSAIL's Speech2Face research paper from IEEE CVPR 2019☆13Mar 25, 2023Updated 3 years ago
- Pytorch implemenation of the model proposed in the paper: Double Multi-Head Attention for Speaker Verification☆19Jul 25, 2024Updated 2 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Simple tools for ComfyUI☆19Jun 27, 2026Updated 2 months ago
- ☆48May 4, 2024Updated 2 years ago
- An simple application to create synchronised audio from a srt file☆17Mar 4, 2025Updated last year
- INA's library with pretrained models for gender and age prediction from faces.☆24Oct 7, 2024Updated last year
- This project aims at giving the best customer service ever using the power of LLM models like GPT.☆10Jun 29, 2023Updated 3 years ago
- Tools for parsing the audio track in television news programs☆19Apr 24, 2021Updated 5 years ago
- The MAVD represents Mandarin Audio-Visual dataset with Depth information. MAVD has a rich variety of modal data, including audio, RGB ima…☆20Apr 22, 2024Updated 2 years ago
- CVPR2025-Multi-party Collaborative Attention Control for Image Customization☆17May 14, 2025Updated last year
- Multi-Stage Face-Voice Association Learning with Keynote Speaker Diarization (ACM MM 2024)☆22Jul 25, 2024Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- This UE4 project contains the Telekinesis Mechanic for Control☆11Jul 26, 2020Updated 6 years ago
- Implementing isometric 3D effect in a pure 2D environment.☆14Apr 21, 2021Updated 5 years ago
- A remaster of Monolith's 1997 Captain Claw game using Unreal Engine 4.26.2 written in C++ and focusses on Object Oriented Design.☆13Dec 16, 2022Updated 3 years ago
- Auto translate all languages to english☆19Sep 9, 2024Updated 2 years ago
- An end-to-end open-source data stack for crawling and visualizing real estate data, facilitating insights into market trends.☆15Mar 16, 2026Updated 6 months ago
- Flan T5 LLM fine-tuning, by attaching a regression model last hidden layers activations. Runs on colab with A100 40gb☆13Mar 24, 2023Updated 3 years ago
- 最新pytorch分布式训练,单机多卡,多机多卡整理(多GPU)☆18Mar 25, 2024Updated 2 years ago