Extract a target speaker’s clean, non-overlapped speech from multi-speaker audio and export word-safe LJSpeech-style TTS datasets.
☆23Jun 14, 2026Updated 3 months ago
Alternatives and similar repositories for Timbre
Users that are interested in Timbre are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Open-weights voice acting pipeline combining zero-shot voice cloning with natural-language direction. Provide a reference voice (or gener…☆18May 25, 2026Updated 4 months ago
- ☆31Feb 14, 2026Updated 7 months ago
- ☆26Mar 1, 2026Updated 7 months ago
- ArtSpeech: Adaptive Text-to-Speech Synthesis with Articulatory Representations☆22Sep 21, 2025Updated last year
- A variable-frame-rate 16 kHz speech codec based on FocalCodec☆21Feb 11, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- High fidelity neural audio codec for TTS models☆37Dec 22, 2025Updated 9 months ago
- ☆28Nov 15, 2023Updated 2 years ago
- An attempt to reproduce CALM (Continuous Audio Language Models) using DACVAE as the audio VAE.☆19Feb 20, 2026Updated 7 months ago
- A highly optimized engine for neutts-air model to generate minutes of audio in seconds. Over 200x realtime on modern hardware!☆119Nov 24, 2025Updated 10 months ago
- Unofficial implementation of ConvNeXt-TTS powered by lightning☆18Oct 20, 2024Updated last year
- [ICASSP 2026] Official code for "Measuring Prosody Diversity in Zero-Shot TTS: A New Metric, Benchmark, and Exploration"☆18Apr 16, 2026Updated 5 months ago
- Echo-TTS inference codebase☆226Dec 5, 2025Updated 9 months ago
- A highly optimized engine for maya-1 tts model to generate minutes of audio in seconds.☆65Nov 17, 2025Updated 10 months ago
- An AR+AR TTS attempt.☆18Jan 13, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Kanade is a single-layer disentangled speech tokenizer that extracts compact tokens suitable for both generative and discriminative model…☆118Jul 18, 2026Updated 2 months ago
- ☆12Nov 7, 2024Updated last year
- All-in-one Speech Transcription☆11Jun 5, 2026Updated 3 months ago
- Non Parallel Voice Conversion based on VITS☆24Mar 31, 2023Updated 3 years ago
- ☆14Jun 23, 2024Updated 2 years ago
- A universal phone recognizer that can transcribe speech in 100+ languages into IPA☆46Updated this week
- ☆14Nov 22, 2022Updated 3 years ago
- Grapheme-to-phoneme tool for corpus conversion, where phonemes match Phoible inventories☆20Apr 10, 2025Updated last year
- Resources that make every language unique☆33Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- source code of EfficientTTS 2☆22Feb 18, 2024Updated 2 years ago
- ☆21Apr 25, 2026Updated 5 months ago
- An upgrade framework for train and validate compare with icefall using Lightning.☆16Mar 26, 2025Updated last year
- StyleTTS 2 Optimized Training Fork☆32Feb 2, 2025Updated last year
- ☆11Sep 5, 2025Updated last year
- Project for HIDING SPEAKER’S SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINE☆15Nov 30, 2022Updated 3 years ago
- ☆59Feb 8, 2026Updated 7 months ago
- Soprano-Factory: Train your own 2000x realtime text-to-speech model☆262Jan 13, 2026Updated 8 months ago
- This is not remotely close to a finished product, and does not intend to nor does this claim to be working fine-tuning code for MaskGCT. …☆13Dec 4, 2024Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- A collection of all our phonemeizers for dataset construction and inference☆32Feb 21, 2025Updated last year
- pytorch model for contexless-phoneme prediction from speech audio☆32Oct 30, 2025Updated 11 months ago
- GitHub repository linked to AnimeBackgroundGAN HuggingFace Space☆10May 24, 2022Updated 4 years ago
- ☆17Dec 12, 2023Updated 2 years ago
- Whisfusion: Parallel ASR Decoding via a Diffusion Transformer☆31Aug 22, 2025Updated last year
- ☆23Jul 22, 2022Updated 4 years ago
- VyvoTTS: LLM-Based Text-to-Speech Training Framework☆262Sep 17, 2026Updated 2 weeks ago