☆90Dec 31, 2025Updated 7 months ago
Alternatives and similar repositories for DisCo-Speech
Users that are interested in DisCo-Speech are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Scaled diffusion transformer for text-to-speech synthesis (DiT + T5Gemma2 conditioning, TorchTitan & Megatron backends, tested up to 1024…☆24Mar 29, 2026Updated 4 months ago
- Inference for the STFT-VAE continuous audio codec (24kHz, 3.125Hz latent)☆43Jul 12, 2026Updated 3 weeks ago
- [ACL 2026 Main] Training, inference, and testing of the SAC speech codec model.☆108Nov 1, 2025Updated 9 months ago
- [ICASSP 2025] AnCoGen: Analysis, Control and Generation of Speech with a Masked Autoencoder☆14Mar 11, 2025Updated last year
- ☆24Nov 16, 2025Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A Lightweight and Streaming Zero-Shot Voice Conversion via Mean Flows☆300Updated this week
- poorman's ar-dit tts☆45Dec 31, 2025Updated 7 months ago
- KVAE-Audio: a continuous full-band audio waveform autoencoder☆102Jul 23, 2026Updated last week
- 5Hz Deep-Compression Speech VAE for AR-Diffusion and CALMs☆57Nov 19, 2025Updated 8 months ago
- Official code for "WavCube: Unifying Speech Representation for Understanding and Generation via Semantic-Acoustic Joint Modeling"☆62Jun 27, 2026Updated last month
- This repository implement a novel zero-shot TTS framework, named Flamed-TTS, focusing on the efficient generation and dynamic pacing in …☆57Aug 9, 2025Updated 11 months ago
- [ICASSP 2026] Task Vector in TTS: Toward Emotionally Expressive Dialectal Speech Synthesis