A fast, local neural text to speech system
β11,283Aug 26, 2025Updated last year
Alternatives and similar repositories for piper
Users that are interested in piper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ46,026Aug 16, 2024Updated 2 years ago
- Fast and local neural text-to-speech engineβ5,614Updated this week
- A multi-voice TTS system trained with an emphasis on qualityβ14,876Nov 19, 2024Updated last year
- StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Modelsβ6,355Aug 10, 2024Updated 2 years ago
- Port of OpenAI's Whisper model in C/C++β53,755Updated this week
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Faster Whisper transcription with CTranslate2β25,456Nov 19, 2025Updated 10 months ago
- π Text-Prompted Generative Audio Modelβ39,275Aug 19, 2024Updated 2 years ago
- eSpeak NG is an open source speech synthesizer that supports more than hundred languages and accents.β6,852Updated this week
- An Open Source text-to-speech system built by inverting Whisper.β4,648Dec 14, 2025Updated 9 months ago
- Inference and training library for high-quality TTS models.β5,592Dec 10, 2024Updated last year
- Instant voice cloning by MIT and MyShell. Audio foundation model.β37,569Apr 19, 2025Updated last year
- C++ library for converting text to phonemes for Piperβ142Jul 10, 2025Updated last year
- LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.β49,156Updated this week
- Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntimeβ¦β14,843Updated this week
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Converts text to speech in realtimeβ4,031Aug 31, 2026Updated 2 weeks ago
- Robust Speech Recognition via Large-Scale Weak Supervisionβ109,332Aug 31, 2026Updated 2 weeks ago
- Foundational model for human-like, expressive TTSβ4,207Jul 30, 2024Updated 2 years ago
- High-quality multi-lingual text-to-speech library by MyShell.ai. Support English, Spanish, French, Chinese, Japanese and Korean.β7,642Dec 24, 2024Updated last year
- LLM inference in C/C++β128,690Updated this week
- Zero-Shot Speech Editing and Text-to-Speech in the Wildβ8,576May 30, 2026Updated 3 months ago
- WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)β24,106Aug 30, 2026Updated 2 weeks ago
- https://hf.co/hexgrad/Kokoro-82Mβ8,890Aug 6, 2025Updated last year
- Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"β15,243Jul 23, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- SOTA Open Source TTSβ32,735Updated this week
- Silero VAD: pre-trained enterprise-grade Voice Activity Detectorβ10,244Updated this week
- Distribute and run LLMs with a single file.β25,991Updated this week
- SoTA open-source TTSβ26,465Jul 21, 2026Updated last month
- AllTalk is based on the Coqui TTS engine, similar to the Coqui_tts extension for Text generation webUI, however supports a variety of advβ¦β2,433Jan 9, 2026Updated 8 months ago
- Open-source desktop app for local LLMs. Text, vision, tool-calling, OpenAI/Anthropic-compatible API. 100% private.β47,685Aug 17, 2026Updated last month
- Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Nodeβ15,136Aug 9, 2026Updated last month
- Towards Human-Sounding Speechβ6,339Dec 5, 2025Updated 9 months ago
- Get up and running with Kimi, GLM, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.β181,209Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Suno AI's Bark model in C/C++ for fast text-to-speech generationβ870Nov 16, 2024Updated last year
- Dockerized OpenAI-compatible wrapper for Kokoro-82M text-to-speech w/multiplatform CPU, AMD, NVIDIA GPU PyTorch; multi-speaker, clone-tunβ¦β5,456Sep 10, 2026Updated last week
- A TTS model capable of generating ultra-realistic dialogue in one pass.β19,399Nov 19, 2025Updated 10 months ago
- A fast local neural text to speech engine for Mycroftβ1,263Mar 25, 2025Updated last year
- Local voice recording for creating Piper datasetsβ221Aug 10, 2026Updated last month
- An open source voice assistant toolkit for many human languagesβ382Dec 26, 2023Updated 2 years ago
- Silero Models: pre-trained text-to-speech models made embarrassingly simpleβ6,115Jul 31, 2026Updated last month