πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
β36Mar 10, 2022Updated 4 years ago
Alternatives and similar repositories for Coqui-TTS
Users that are interested in Coqui-TTS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ16Updated this week
- β22Apr 4, 2023Updated 3 years ago
- [TOMM 2024] Automatic Lyric Transcription and Automatic Music Transcription from Multimodal Singingβ28Aug 30, 2024Updated 2 years ago
- Pythonηι³ι’ε·₯ε ·β16Dec 5, 2025Updated 9 months ago
- Official implementation of "Automatic Tuning of Loss Trade-offs without Hyper-parameter Search in End-to-End Zero-Shot Speech Synthesis",β¦β80May 29, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Ultrafast GAN based Vocoder for Text to Speechβ50Jul 16, 2022Updated 4 years ago
- β24Mar 15, 2022Updated 4 years ago
- Unofficial Pytorch implementation of SNAC: Speaker-normalized affine coupling layer in flow-based architecture for zero-shot multi-speakeβ¦β57Aug 7, 2023Updated 3 years ago
- pytorch implementation for MultiSpeech: Multi-Speaker Text to Speech with Transformer paperβ21Jun 23, 2022Updated 4 years ago
- The code for aishell-3 baseline acoustic modelβ70Nov 30, 2020Updated 5 years ago
- Official repository of https://doi.org/10.1109/TASLP.2022.3167258. More up-to-date code is in "refactor" branch.β192Jun 8, 2023Updated 3 years ago
- GitHub repository for AudioToolAgentβ21Feb 13, 2026Updated 7 months ago
- β26Sep 22, 2022Updated 4 years ago
- YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyoneβ1,052Nov 4, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The source code of our paper "Diffsound: discrete diffusion model for text-to-sound generation"β366Aug 3, 2023Updated 3 years ago
- Researchers who published code, models (in some cases), and demo apps (in few cases) along with their SOTA paperβ12Oct 19, 2023Updated 2 years ago
- Autovocoder: Fast Waveform Generation from a Learned Speech Representation using Differentiable Digital Signal Processingβ71Dec 2, 2022Updated 3 years ago
- This is Pytorch Implementation of Google's Non-attentive Tacotron.β57Dec 21, 2022Updated 3 years ago
- The source code of Tim-TSENetβ15Apr 22, 2022Updated 4 years ago
- torch version of LPCNetβ22Jul 8, 2020Updated 6 years ago
- β13Sep 20, 2023Updated 3 years ago
- The open source code for SimpleSpeech seriesβ146Oct 8, 2024Updated last year
- GE2E Speaker Encoder - Generalized End-To-End Loss for Speaker Verificationβ14May 17, 2020Updated 6 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- β63Jan 15, 2024Updated 2 years ago
- S3PRL-VC: A Voice Conversion Toolkit based on S3PRLβ101Mar 15, 2026Updated 6 months ago
- Yet Another Speech Parser (YASP), uses pocketsphinx to parse speech and output a JSON file describing the timingβ16Jun 21, 2020Updated 6 years ago
- This repository provides UNOFFICIAL Bunched LPCNet implementations with Pytorch.β14Jun 17, 2021Updated 5 years ago
- Linear Prediction Coefficients estimation from mel-spectrogram implemented in Python based on Levinson-Durbin algorithm.β72Mar 19, 2021Updated 5 years ago
- TTS Text Analyzerβ31Jul 20, 2023Updated 3 years ago
- β23Sep 14, 2021Updated 5 years ago
- PyTorch Implementation of NCSOFT's FastPitchFormant: Source-filter based Decomposed Modeling for Speech Synthesisβ74Aug 3, 2021Updated 5 years ago
- An unofficial PyTorch implementation of "HiFi-GAN: High-Fidelity Denoising and Dereverberation Based on Speech Deep Features in Adversariβ¦β25Feb 5, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PyTorch implementation of Listen, Attend and Spell (LAS) speech recognition paperβ12Mar 4, 2022Updated 4 years ago
- Coqui STT (πΈSTT) based forced alignment toolβ13Feb 24, 2022Updated 4 years ago
- Pytorch implementation of Tacotron, a speech synthesis end-to-end generative TTS model.β29Mar 14, 2019Updated 7 years ago
- β16Sep 6, 2021Updated 5 years ago
- Implementation of Spear-TTS - multi-speaker text-to-speech attention network, in Pytorchβ278Oct 30, 2023Updated 2 years ago
- π The chrome extension to make learning from YouTube faster & easier.β11Jan 9, 2022Updated 4 years ago
- β19Jul 12, 2020Updated 6 years ago