The Smallest English TTS Model with only 1M parameters
☆619Apr 10, 2026Updated 6 months ago
Alternatives and similar repositories for tiny-tts
Users that are interested in tiny-tts are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Fast inference engine for DACVAE, a neural audio codec that compresses and reconstructs audio using a convolutional encoder-decoder with …☆21Mar 17, 2026Updated 6 months ago
- High fidelity neural audio codec for TTS models☆37Dec 22, 2025Updated 9 months ago
- High quality text-to-speech based on StyleTTS 2.☆79Apr 6, 2026Updated 6 months ago
- Soprano-Factory: Train your own 2000x realtime text-to-speech model☆260Jan 13, 2026Updated 8 months ago
- Building actual open source Multilingual speech model including dataset.☆58Oct 1, 2026Updated last week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Kanade is a single-layer disentangled speech tokenizer that extracts compact tokens suitable for both generative and discriminative model…☆119Jul 18, 2026Updated 2 months ago
- Open-source text-to-speech model from KRAFTON trained exclusively on public speech data, with curated datasets and reproducible training …☆101May 21, 2026Updated 4 months ago
- A universal phone recognizer that can transcribe speech in 100+ languages into IPA☆46Sep 30, 2026Updated last week
- ☆375Aug 28, 2025Updated last year
- Unofficial implementation of ConvNeXt-TTS powered by lightning☆18Oct 20, 2024Updated last year
- ru-normalizr — лучший нормализатор русского текста без LLM. Приводит числа, даты, время, сокращения, римские цифры, символы и латиницу в …☆26Aug 7, 2026Updated 2 months ago
- A highly compressive and high-quality neural audio codec for speech models.☆275Jan 23, 2026Updated 8 months ago
- Echo-TTS inference codebase☆226Dec 5, 2025Updated 10 months ago
- A lightweight, efficient variation of the StyleTTS 2 text‐to‐speech model.☆53May 22, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Inference and deployment toolkit for Svara-TTS, an open-source multilingual text-to-speech model for Indic languages☆32Apr 1, 2026Updated 6 months ago
- Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.☆13,775Sep 9, 2026Updated last month
- Research-first architecture engine for Python, TS, Go and Rust. Mines GitHub Issues for real production failures before scaffolding, then…☆49Oct 3, 2026Updated last week
- ☆12Nov 7, 2024Updated last year
- superfast text to speech in any voice☆63Feb 16, 2026Updated 7 months ago
- IPA Phonemizer/Dephonemizer for 140 human languages☆64Oct 2, 2026Updated last week
- OLaPh (Optimal Language Phonemizer) is a multilingual phonemization framework that converts text into phonemes surpassing the quality of …☆23Aug 19, 2026Updated last month
- Based on the implementation of Google's TurboQuant (ICLR 2026) — Quansloth brings elite KV cache compression to local LLM inference. Qua…☆155May 13, 2026Updated 4 months ago
- [EMNLP 2024] ESC: Efficient Speech Coding with Cross-Scale Residual Vector Quantized Transformers☆127Mar 20, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- An experiment to use a webcam as a game input device.☆12Nov 22, 2022Updated 3 years ago
- Conformer block with Rotary Position Embedding, modified from lucidrains' implement☆19Sep 13, 2024Updated 2 years ago
- Training code for kokoro tts model☆46Nov 15, 2025Updated 10 months ago
- A TTS that fits in your CPU (and pocket)☆9,839Updated this week
- A lightweight text-to-speech model with zero-shot voice cloning☆970Sep 1, 2026Updated last month
- Project for HIDING SPEAKER’S SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINE☆15Nov 30, 2022Updated 3 years ago
- A highly optimized engine for neutts-air model to generate minutes of audio in seconds. Over 200x realtime on modern hardware!☆119Nov 24, 2025Updated 10 months ago
- Arabic Grapheme-to-Phoneme (G2P) Conversion☆16Mar 15, 2025Updated last year
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 🌋LavaSR: Fast Speech restoration and enhancement☆610Jun 19, 2026Updated 3 months ago
- [TAFFC 2025] The official implementation of EmoSphere++: Emotion-Controllable Zero-Shot Text-to-Speech via Emotion-Adaptive Spherical Vec…☆132Jul 16, 2026Updated 2 months ago
- Labeled data for homograph disambiguation☆63Jun 1, 2023Updated 3 years ago
- Your AI colleague, in the apps you already use.☆14Updated this week
- An espeak-compatible, permissively-licensed IPA phonemizer (G2P) based on DeepPhonemizer. Usable as a drop-in replacement for espeak's GP…☆112Mar 15, 2026Updated 6 months ago
- VellumForge2 is a Golang CLI for generating high-quality Direct Preference Optimization datasets via a hierarchical prompt pipeline with …☆22Feb 15, 2026Updated 7 months ago
- ☆22Aug 21, 2025Updated last year