Extract a target speaker’s clean, non-overlapped speech from multi-speaker audio and export word-safe LJSpeech-style TTS datasets.
☆23Jun 14, 2026Updated 2 months ago
Alternatives and similar repositories for Timbre
Users that are interested in Timbre are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Open-weights voice acting pipeline combining zero-shot voice cloning with natural-language direction. Provide a reference voice (or gener…☆18May 25, 2026Updated 3 months ago
- Distributed workflow automation engine built for AI-native workloads.☆18Updated this week
- ☆29Feb 14, 2026Updated 6 months ago
- ☆24Mar 1, 2026Updated 6 months ago
- ArtSpeech: Adaptive Text-to-Speech Synthesis with Articulatory Representations☆22Sep 21, 2025Updated 11 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A variable-frame-rate 16 kHz speech codec based on FocalCodec☆21Feb 11, 2026Updated 7 months ago
- High fidelity neural audio codec for TTS models☆36Dec 22, 2025Updated 8 months ago
- ☆28Nov 15, 2023Updated 2 years ago
- An attempt to reproduce CALM (Continuous Audio Language Models) using DACVAE as the audio VAE.☆19Feb 20, 2026Updated 6 months ago
- A highly optimized engine for neutts-air model to generate minutes of audio in seconds. Over 200x realtime on modern hardware!☆120Nov 24, 2025Updated 9 months ago
- Unofficial implementation of ConvNeXt-TTS powered by lightning☆18Oct 20, 2024Updated last year
- [ICASSP 2026] Official code for "Measuring Prosody Diversity in Zero-Shot TTS: A New Metric, Benchmark, and Exploration"☆18Apr 16, 2026Updated 4 months ago
- Echo-TTS inference codebase☆224Dec 5, 2025Updated 9 months ago
- A highly optimized engine for maya-1 tts model to generate minutes of audio in seconds.☆66Nov 17, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- An AR+AR TTS attempt.☆18Jan 13, 2025Updated last year
- Kanade is a single-layer disentangled speech tokenizer that extracts compact tokens suitable for both generative and discriminative model…☆115Jul 18, 2026Updated last month
- A universal phone recognizer that can transcribe speech in 70+ languages into IPA☆40Jun 9, 2026Updated 3 months ago
- ☆12Nov 7, 2024Updated last year
- Grapheme-to-phoneme tool for corpus conversion, where phonemes match Phoible inventories☆19Apr 10, 2025Updated last year
- All-in-one Speech Transcription☆11Jun 5, 2026Updated 3 months ago
- Non Parallel Voice Conversion based on VITS☆24Mar 31, 2023Updated 3 years ago
- ☆15Jun 23, 2024Updated 2 years ago
- ☆14Nov 22, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Resources that make every language unique☆33Aug 25, 2026Updated 2 weeks ago
- source code of EfficientTTS 2☆22Feb 18, 2024Updated 2 years ago
- An upgrade framework for train and validate compare with icefall using Lightning.☆16Mar 26, 2025Updated last year
- ☆21Apr 25, 2026Updated 4 months ago
- StyleTTS 2 Optimized Training Fork☆32Feb 2, 2025Updated last year
- ☆11Sep 5, 2025Updated last year
- Project for HIDING SPEAKER’S SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINE☆15Nov 30, 2022Updated 3 years ago
- ☆59Feb 8, 2026Updated 7 months ago
- Soprano-Factory: Train your own 2000x realtime text-to-speech model☆260Jan 13, 2026Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This is not remotely close to a finished product, and does not intend to nor does this claim to be working fine-tuning code for MaskGCT. …☆13Dec 4, 2024Updated last year
- pytorch model for contexless-phoneme prediction from speech audio☆32Oct 30, 2025Updated 10 months ago
- A collection of all our phonemeizers for dataset construction and inference☆32Feb 21, 2025Updated last year
- GitHub repository linked to AnimeBackgroundGAN HuggingFace Space☆10May 24, 2022Updated 4 years ago
- ☆17Dec 12, 2023Updated 2 years ago
- Whisfusion: Parallel ASR Decoding via a Diffusion Transformer☆31Aug 22, 2025Updated last year
- ☆23Jul 22, 2022Updated 4 years ago