Easy fine-tuning for Qwen3-TTS: Fast voice cloning and high-quality multilingual speech synthesis.
☆115May 29, 2026Updated 2 months ago
Alternatives and similar repositories for Qwen3-TTS-EasyFinetuning
Users that are interested in Qwen3-TTS-EasyFinetuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- One-command fine-tuning for Qwen3-TTS text-to-speech model with custom voice samples☆29Jan 28, 2026Updated 6 months ago
- LoRA-based phoneme/prosody control for LLM-based TTS with no G2P - Lightweight adapter for edit and control the target language's phoneme…☆26Jul 8, 2026Updated 3 weeks ago
- speaker-disentangled speech linguistic content quantizer☆26Mar 19, 2025Updated last year
- [ICASSP 2026] Task Vector in TTS: Toward Emotionally Expressive Dialectal Speech Synthesis☆40Dec 24, 2025Updated 7 months ago
- [ICLR2026] FlexiCodec: A Dynamic Neural Audio Codec for Low Frame Rates☆51Jul 1, 2026Updated 3 weeks ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into one☆26Aug 5, 2024Updated last year
- An All-in-One Speech, Sound, Music Codec with Single Nested Codebook☆28Oct 11, 2025Updated 9 months ago
- [ACM MM 2025] AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation☆24Oct 28, 2025Updated 9 months ago
- Collection of scripts from mHuBERT-147.☆35Nov 19, 2024Updated last year
- Official code for"DiaMoE-TTS: A Unified IPA-based Dialect TTS Framework with Mixture-of-Experts and Parameter-Efficient Zero-Shot Adaptat…☆246Nov 28, 2025Updated 8 months ago
- The demo page for ALMTokenizer☆59Apr 14, 2025Updated last year
- SpeechJudge: Towards Human-Level Judgment for Speech Naturalness (https://arxiv.org/abs/2511.07931)☆79Dec 23, 2025Updated 7 months ago
- IndexTTS Fine-tuning notebooks☆139Jun 17, 2025Updated last year
- CosyVoice_DPO_NOTES: Supercharge Your Cosyvoice model with Cutting-Edge DPO Fine-Tuning!☆126Aug 8, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is the official train-dev-test release of the Interspeech2024 Discrete Speech Representation Challenge.☆32Jan 26, 2024Updated 2 years ago
- [ICASSP2025] Official code for VoiceDiT: Dual-Condition Diffusion Transformer for Environment-Aware Speech Synthesis☆52Apr 9, 2025Updated last year
- ☆16Nov 11, 2024Updated last year
- Convert English text from written expressions into spoken forms☆32Jun 22, 2022Updated 4 years ago
- [ACMMM'2024] Generative Expressive Conversational Speech Synthesis☆45Oct 28, 2024Updated last year
- Codebase for 'ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining'☆23Jun 20, 2026Updated last month
- PitchVC: Pitch Conditioned Any-to-Many Voice Conversion☆35Jun 6, 2024Updated 2 years ago
- [ACL 2026 Main] Training, inference, and testing of the SAC speech codec model.☆108Nov 1, 2025Updated 8 months ago
- Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale☆29Aug 4, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- FINALLY: Fast and universal speech enhancement model delivering studio-quality audio for a wide range of recordings.☆28Apr 1, 2026Updated 3 months ago
- PyTorch Implementation of [WMCodec: End-to-End Neural Speech Codec with Deep Watermarking for Authenticity Verification](https://arxiv.or…☆18Jul 31, 2025Updated 11 months ago
- ☆41May 15, 2023Updated 3 years ago
- Real-time streaming TTS for Qwen3-TTS with two-phase latency, Hann crossfade, and torch.compile + CUDA graphs optimizations☆93Feb 21, 2026Updated 5 months ago
- [Interspeech 2025] Official implementation of "Training-Free Voice Conversion with Factorized Optimal Transport"☆45Sep 24, 2025Updated 10 months ago
- UTokyo-SaruLab MOS Prediction System☆357Apr 2, 2026Updated 3 months ago
- SpeechGLUE is a speech version of the GLUE benchmark, driven by text-to-speech.☆13Jun 2, 2023Updated 3 years ago
- Fork with streaming inference support + ~6× faster inference☆256Jun 3, 2026Updated last month
- ☆11Aug 1, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆10Apr 17, 2024Updated 2 years ago
- Conformer block with Rotary Position Embedding, modified from lucidrains' implement☆19Sep 13, 2024Updated last year
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆17Apr 15, 2026Updated 3 months ago
- Speech samples and code of BEdit-TTS☆34Oct 8, 2023Updated 2 years ago
- poorman's ar-dit tts☆45Dec 31, 2025Updated 6 months ago
- Real-time text-to-speech with Qwen3-TTS☆1,261Jul 17, 2026Updated last week
- Open-weights voice acting pipeline combining zero-shot voice cloning with natural-language direction. Provide a reference voice (or gener…☆17May 25, 2026Updated 2 months ago