πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
β37Mar 10, 2022Updated 4 years ago
Alternatives and similar repositories for Coqui-TTS
Users that are interested in Coqui-TTS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- β22Apr 4, 2023Updated 3 years ago
- [TOMM 2024] Automatic Lyric Transcription and Automatic Music Transcription from Multimodal Singingβ28Aug 30, 2024Updated last year
- Pythonηι³ι’ε·₯ε ·β16Dec 5, 2025Updated 8 months ago
- Official implementation of "Automatic Tuning of Loss Trade-offs without Hyper-parameter Search in End-to-End Zero-Shot Speech Synthesis",β¦β80May 29, 2023Updated 3 years ago
- Ultrafast GAN based Vocoder for Text to Speechβ50Jul 16, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- β24Mar 15, 2022Updated 4 years ago
- Unofficial Pytorch implementation of SNAC: Speaker-normalized affine coupling layer in flow-based architecture for zero-shot multi-speakeβ¦β57Aug 7, 2023Updated 3 years ago
- pytorch implementation for MultiSpeech: Multi-Speaker Text to Speech with Transformer paperβ21Jun 23, 2022Updated 4 years ago
- The code for aishell-3 baseline acoustic modelβ70Nov 30, 2020Updated 5 years ago
- Official repository of https://doi.org/10.1109/TASLP.2022.3167258. More up-to-date code is in "refactor" branch.β192Jun 8, 2023Updated 3 years ago
- GitHub repository for AudioToolAgentβ20Feb 13, 2026Updated 6 months ago
- β26Sep 22, 2022Updated 3 years ago
- YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyoneβ1,053Nov 4, 2024Updated last year
- The source code of our paper "Diffsound: discrete diffusion model for text-to-sound generation"β366Aug 3, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Researchers who published code, models (in some cases), and demo apps (in few cases) along with their SOTA paperβ12Oct 19, 2023Updated 2 years ago
- Autovocoder: Fast Waveform Generation from a Learned Speech Representation using Differentiable Digital Signal Processingβ71Dec 2, 2022Updated 3 years ago
- This is Pytorch Implementation of Google's Non-attentive Tacotron.β57Dec 21, 2022Updated 3 years ago
- The source code of Tim-TSENetβ15Apr 22, 2022Updated 4 years ago
- torch version of LPCNetβ22Jul 8, 2020Updated 6 years ago
- β13Sep 20, 2023Updated 2 years ago
- The open source code for SimpleSpeech seriesβ147Oct 8, 2024Updated last year
- GE2E Speaker Encoder - Generalized End-To-End Loss for Speaker Verificationβ14May 17, 2020Updated 6 years ago
- β63Jan 15, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- S3PRL-VC: A Voice Conversion Toolkit based on S3PRLβ101Mar 15, 2026Updated 5 months ago
- Yet Another Speech Parser (YASP), uses pocketsphinx to parse speech and output a JSON file describing the timingβ16Jun 21, 2020Updated 6 years ago
- This repository provides UNOFFICIAL Bunched LPCNet implementations with Pytorch.β14Jun 17, 2021Updated 5 years ago
- Linear Prediction Coefficients estimation from mel-spectrogram implemented in Python based on Levinson-Durbin algorithm.β72Mar 19, 2021Updated 5 years ago
- TTS Text Analyzerβ31Jul 20, 2023Updated 3 years ago
- β23Sep 14, 2021Updated 4 years ago
- PyTorch Implementation of NCSOFT's FastPitchFormant: Source-filter based Decomposed Modeling for Speech Synthesisβ74Aug 3, 2021Updated 5 years ago
- An unofficial PyTorch implementation of "HiFi-GAN: High-Fidelity Denoising and Dereverberation Based on Speech Deep Features in Adversariβ¦β25Feb 5, 2021Updated 5 years ago
- PyTorch implementation of Listen, Attend and Spell (LAS) speech recognition paperβ12Mar 4, 2022Updated 4 years ago
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Coqui STT (πΈSTT) based forced alignment toolβ13Feb 24, 2022Updated 4 years ago
- Pytorch implementation of Tacotron, a speech synthesis end-to-end generative TTS model.β29Mar 14, 2019Updated 7 years ago
- β15Sep 6, 2021Updated 4 years ago
- Implementation of Spear-TTS - multi-speaker text-to-speech attention network, in Pytorchβ277Oct 30, 2023Updated 2 years ago
- β19Jul 12, 2020Updated 6 years ago
- π The chrome extension to make learning from YouTube faster & easier.β11Jan 9, 2022Updated 4 years ago
- β19Feb 2, 2023Updated 3 years ago