This project explores zero-shot emotional speech synthesis using EMOD, a novel approach combining emotion and content embeddings for multilingual and cross-lingual emotion transfer. Built on a VITS-based TTS model, it preserves speaker identity while enhancing expressiveness, enabling emotion transfer across languages and genders efficiently.
☆19Jun 26, 2026Updated 2 months ago
Alternatives and similar repositories for Emotion-TTS-Emebddings
Users that are interested in Emotion-TTS-Emebddings are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Aug 19, 2024Updated 2 years ago
- Official code for ProtoDiff☆16Dec 6, 2023Updated 2 years ago
- ☆123Oct 24, 2022Updated 3 years ago
- Please visit: https://thuhcsi.github.io/icassp2021-emotion-tts/☆34Mar 17, 2023Updated 3 years ago
- Deformable Convolutional Networks v2 with Pytorch☆10Jul 29, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The electronic Holly Quran browser Elforkane☆11Nov 14, 2021Updated 4 years ago
- ☆11Mar 28, 2021Updated 5 years ago
- ☆18Sep 19, 2023Updated 3 years ago
- Embedded Tajweed annotation for the Qur'an☆12Nov 30, 2025Updated 9 months ago
- S3PRL for Speech Emotion Recognition (see s3prl > downstream)☆15Feb 28, 2026Updated 6 months ago
- ☆11Jan 22, 2017Updated 9 years ago
- Y-vector: Multiscale Waveform Encoder for Speaker Embedding☆24Jul 16, 2024Updated 2 years ago
- Learning Transferable Features with Deep Adaptation Networks☆12Jul 18, 2023Updated 3 years ago
- ZET-Speech: Zero-shot adaptive Emotion-controllable Text-to-Speech Synthesis with Diffusion and Style-based Models (TTS)☆10Mar 9, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆15Jun 26, 2023Updated 3 years ago
- ☆20Aug 23, 2024Updated 2 years ago
- [APSIPA'22] Exploring Speaker Age Estimation on Different Self-Supervised Learning Models☆14Oct 19, 2022Updated 3 years ago
- [ACMMM'2024] Generative Expressive Conversational Speech Synthesis☆45Oct 28, 2024Updated last year
- ☆57Jul 16, 2025Updated last year
- A tool for visualizing emotions in music using a Python wrapper for Spotify API. Independent post-baccalaureate research by Nick Stapleto…☆13Jun 2, 2023Updated 3 years ago
- ☆12Oct 28, 2023Updated 2 years ago
- Release code for light-weight calibrator: a separable component for unsupervised domain adaptation☆13Jul 17, 2021Updated 5 years ago
- state-of-the-art models for diacritics restoration for Arabic language☆16Feb 23, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Text to Speech with PyTorch (English and Mongolian)☆13May 3, 2020Updated 6 years ago
- [INTERSPEECH 2024] The official implementation of EmoSphere-TTS: Emotional Style and Intensity Modeling via Spherical Emotion Vector for …☆186Jul 16, 2026Updated 2 months ago
- Awesome list of TTS papers with audio samples☆60Aug 18, 2020Updated 6 years ago
- A pep8 and pyflakes checker for Gedit☆18Nov 12, 2017Updated 8 years ago
- Awesome TTS☆63Sep 16, 2021Updated 5 years ago
- Pytorch 1.0 codes(including cuda codes) for Deformable Convolution Version 2☆18Mar 2, 2019Updated 7 years ago
- 模式识别期末项目-基于Keras的人物面部表情识别☆11Jun 25, 2019Updated 7 years ago
- A python toolkit to detect, segment, and count coughs☆16Aug 1, 2026Updated last month
- unofficial pytorch implementation of HiFi-GAN with fast MISR.☆15Mar 21, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 基于AWS IoT构建车联网平台的演示环境☆17Nov 27, 2019Updated 6 years ago
- Provably Secure Steganography☆22Sep 13, 2025Updated last year
- This is the implementation for "ControlVC: Zero-Shot Voice Conversion with Time-Varying Controls on Pitch and Rhythm"☆132Nov 29, 2023Updated 2 years ago
- Matlab tools for pathological voice analysis☆14May 12, 2023Updated 3 years ago
- [IEEE, TASLP, 2023] The code of the paper "Multi-Source Discriminant Subspace Alignment for Cross-Domain Speech Emotion Recognition".☆19Sep 27, 2024Updated last year
- Transferability of cross-lingual and cross-age speech emotion recognition☆21Jun 30, 2023Updated 3 years ago
- 车牌识别,参考https://github.com/wzh191920/License-Plate-Recognition☆16Sep 26, 2018Updated 7 years ago