Multilingual G2P in 100 languages
☆396May 26, 2023Updated 3 years ago
Alternatives and similar repositories for CharsiuG2P
Users that are interested in CharsiuG2P are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Charsiu: A neural phonetic aligner.☆351Sep 19, 2022Updated 4 years ago
- Grapheme to phoneme conversion with deep learning.☆435Dec 8, 2023Updated 2 years ago
- phoneme tokenizer and grapheme-to-phoneme model for 8k languages☆176Jun 9, 2023Updated 3 years ago
- XPhoneBERT: A Pre-trained Multilingual Model for Phoneme Representations for Text-to-Speech (INTERSPEECH 2023)☆358Jul 22, 2024Updated 2 years ago
- Official implementation of EdiTTS: Score-based Editing for Controllable Text-to-Speech (INTERSPEECH 2022)☆122Jan 24, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Official implementation of the source-filter HiFiGAN vocoder☆278Jul 29, 2023Updated 3 years ago
- Official repository of DailyTalk: Spoken Dialogue Dataset for Conversational Text-to-Speech, ICASSP 2023☆262Jun 5, 2025Updated last year
- HiFTNet: A Fast High-Quality Neural Vocoder with Harmonic-plus-Noise Filter and Inverse Short Time Fourier Transform☆260Jan 14, 2025Updated last year
- [ICASSP 2024] This is the official code for "VoiceFlow: Efficient Text-to-Speech with Rectified Flow Matching"☆375Sep 3, 2024Updated 2 years ago
- Avocodo: Generative Adversarial Network for Artifact-free Vocoder☆122Jul 14, 2022Updated 4 years ago
- Phoneme-Level BERT for Enhanced Prosody of Text-to-Speech with Grapheme Predictions☆269Jan 13, 2025Updated last year
- A Neural Grapheme-to-Phoneme Conversion Package for Mandarin Chinese Based on a New Open Benchmark Dataset☆369Dec 24, 2021Updated 4 years ago
- Pytorch implementation of BigVSAN☆203Dec 9, 2025Updated 9 months ago
- An espeak-compatible, permissively-licensed IPA phonemizer (G2P) based on DeepPhonemizer. Usable as a drop-in replacement for espeak's GP…☆112Mar 15, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A differentiable version of SPTK☆202Sep 16, 2026Updated last week
- Keyword spotting and forced alignment in any language☆102Jun 15, 2026Updated 3 months ago
- ☆88Nov 1, 2022Updated 3 years ago
- Official implementation of "Avocodo: Generative Adversarial Network for Artifact-Free Vocoder" (AAAI2023)☆154Feb 1, 2023Updated 3 years ago
- Train the next generation of TTS systems.☆169Sep 13, 2024Updated 2 years ago
- ICASSP 2023 Accepted☆191May 6, 2024Updated 2 years ago
- An official implementation of "UnitSpeech: Speaker-adaptive Speech Synthesis with Untranscribed Data"☆136Aug 17, 2023Updated 3 years ago
- Provides training, inference and voice conversion recipes for RADTTS and RADTTS++: Flow-based TTS models with Robust Alignment Learning, …☆292Apr 6, 2023Updated 3 years ago
- Official implementation of SawSing (ISMIR'22)☆275Aug 28, 2022Updated 4 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- An Open-source Streaming High-fidelity Neural Audio Codec☆515Mar 4, 2025Updated last year
- ☆202Updated this week
- ☆161Sep 19, 2022Updated 4 years ago
- ☆260May 15, 2023Updated 3 years ago
- pytorch model for contexless-phoneme prediction from speech audio☆32Oct 30, 2025Updated 10 months ago
- Grapheme-to-Phoneme transductions that preserve input and output indices, and support cross-lingual g2p!☆203Updated this week
- PyTorch Implementation of ProDiff (ACM-MM'22) with a Extremely-Fast diffusion speech synthesis pipeline☆432Apr 19, 2023Updated 3 years ago
- ☆47Apr 16, 2023Updated 3 years ago
- This is the source code of the paper "Neural grapheme-to-phoneme conversion with pretrained grapheme models☆48Mar 25, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis☆1,161Aug 29, 2026Updated 3 weeks ago
- ☆53Aug 28, 2024Updated 2 years ago
- multilingual speech aligner☆79Nov 19, 2023Updated 2 years ago
- Simple text to phones converter for multiple languages☆1,571Aug 4, 2026Updated last month
- Code for vec2wav 2.0, a speech token vocoder for VC. Paper: https://arxiv.org/abs/2409.01995☆79Dec 3, 2024Updated last year
- [IJCAI'23] Learning to Speak from Text for Low-Resource TTS☆65May 30, 2023Updated 3 years ago
- The YouTube Text-To-Speech dataset is comprised of waveform audio extracted from YouTube videos alongside their English transcriptions☆53Apr 1, 2021Updated 5 years ago