pytorch model for contexless-phoneme prediction from speech audio
☆32Oct 30, 2025Updated 8 months ago
Alternatives and similar repositories for contexless-phonemes-CUPE
Users that are interested in contexless-phonemes-CUPE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Extract phoneme-level timestamps from speeh audio.☆156Jun 7, 2026Updated last month
- This repository implement a novel zero-shot TTS framework, named Flamed-TTS, focusing on the efficient generation and dynamic pacing in …☆57Aug 9, 2025Updated 11 months ago
- Inference for the STFT-VAE continuous audio codec (24kHz, 3.125Hz latent)☆43Jul 12, 2026Updated 2 weeks ago
- ☆82Jan 22, 2025Updated last year
- ☆16Nov 11, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ACL 2025] OZSpeech: One-step Zero-shot Speech Synthesis with Learned-Prior-Conditioned Flow Matching☆45Feb 9, 2025Updated last year
- This is the official repository of ``Scalable Neural Vocoder from Range-Null Space Decomposition'', which is submitted to TPAMI.☆54Oct 11, 2025Updated 9 months ago
- ☆56Jul 16, 2025Updated last year
- ☆61Nov 4, 2023Updated 2 years ago
- ☆19Mar 22, 2024Updated 2 years ago
- PyTorch implementation of Miipher-2 [2025] which is a speech restoration model by Google DeepMind☆70Sep 22, 2025Updated 10 months ago
- ☆25Mar 6, 2024Updated 2 years ago
- Try to replicate the architecture of MiniMaxTTS mentioned in it's technical report☆47Sep 2, 2025Updated 10 months ago
- ☆14Jun 16, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆27Jun 22, 2026Updated last month
- High quality text-to-speech based on StyleTTS 2.☆78Apr 6, 2026Updated 3 months ago
- Viterbi decoding in PyTorch☆42May 5, 2026Updated 2 months ago
- PromptTTS++: Controlling Speaker Identity in Prompt-Based Text-To-Speech Using Natural Language Descriptions☆86Oct 11, 2024Updated last year
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into one☆26Aug 5, 2024Updated last year
- superfast text to speech in any voice☆62Feb 16, 2026Updated 5 months ago
- ☆15Nov 26, 2024Updated last year
- Code for Latent Speech-Text Transformer (LST)☆35Mar 12, 2026Updated 4 months ago
- 🎙️ Automatically transcribe audio/video into high-quality, speaker-specific Text-To-Speech datasets☆142Aug 10, 2025Updated 11 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- An espeak-compatible, permissively-licensed IPA phonemizer (G2P) based on DeepPhonemizer. Usable as a drop-in replacement for espeak's GP…☆111Mar 15, 2026Updated 4 months ago
- Prosody and Pronunciation Modification Network☆64May 5, 2025Updated last year
- ☆17Jun 2, 2025Updated last year
- Variable Bitrate Residual Vector Quantization for Audio Coding☆54May 1, 2025Updated last year
- A lightweight audio codec based on a single quantizer☆72Aug 15, 2025Updated 11 months ago
- This is the repository for the work "BridgeVoC: Revitalizing Neural Vocoder from a Restoration Perspective".☆67Nov 5, 2025Updated 8 months ago
- This is the code for paper: XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs☆97Sep 19, 2025Updated 10 months ago
- ☆35Oct 23, 2025Updated 9 months ago
- Just another FastSpeech 2 but cleaner code :)☆29Jun 28, 2024Updated 2 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆152Apr 25, 2025Updated last year
- poorman's ar-dit tts☆45Dec 31, 2025Updated 6 months ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- Streamable Text-to-Speech model using a language modeling approach, without vector quantization☆108May 20, 2025Updated last year
- Official implementation of the paper "Laughter Synthesis using Pseudo Phonetic Tokens with a Large-scale In-the-wild Laughter Corpus" acc…☆77Jul 16, 2023Updated 3 years ago
- Official code for paper:"Speaking Clearly: A Simplified Whisper-Based Codec for Low-Bitrate Speech Coding"☆37Jan 28, 2026Updated 5 months ago
- ☆28Nov 15, 2023Updated 2 years ago