A fast, local neural text to speech system
β11,276Aug 26, 2025Updated 11 months ago
Alternatives and similar repositories for piper
Users that are interested in piper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ45,869Aug 16, 2024Updated last year
- Fast and local neural text-to-speech engineβ5,078Updated this week
- A multi-voice TTS system trained with an emphasis on qualityβ14,867Nov 19, 2024Updated last year
- StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Modelsβ6,328Aug 10, 2024Updated 2 years ago
- Port of OpenAI's Whisper model in C/C++β52,736Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Faster Whisper transcription with CTranslate2β24,824Nov 19, 2025Updated 8 months ago
- π Text-Prompted Generative Audio Modelβ39,235Aug 19, 2024Updated last year
- eSpeak NG is an open source speech synthesizer that supports more than hundred languages and accents.β6,723Aug 3, 2026Updated last week
- An Open Source text-to-speech system built by inverting Whisper.β4,629Dec 14, 2025Updated 7 months ago
- Inference and training library for high-quality TTS models.β5,588Dec 10, 2024Updated last year
- Instant voice cloning by MIT and MyShell. Audio foundation model.β37,110Apr 19, 2025Updated last year
- C++ library for converting text to phonemes for Piperβ142Jul 10, 2025Updated last year
- LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.β48,349Updated this week
- Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntimeβ¦β14,058Updated this week
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Converts text to speech in realtimeβ4,005Aug 2, 2026Updated last week
- Robust Speech Recognition via Large-Scale Weak Supervisionβ106,961Jul 28, 2026Updated last week
- Foundational model for human-like, expressive TTSβ4,204Jul 30, 2024Updated 2 years ago
- High-quality multi-lingual text-to-speech library by MyShell.ai. Support English, Spanish, French, Chinese, Japanese and Korean.β7,568Dec 24, 2024Updated last year
- LLM inference in C/C++β123,193Updated this week
- Zero-Shot Speech Editing and Text-to-Speech in the Wildβ8,565May 30, 2026Updated 2 months ago
- WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)β23,499Jul 13, 2026Updated 3 weeks ago
- https://hf.co/hexgrad/Kokoro-82Mβ8,344Aug 6, 2025Updated last year
- Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"β15,090Jul 23, 2026Updated 2 weeks ago
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- SOTA Open Source TTSβ32,111Updated this week
- Silero VAD: pre-trained enterprise-grade Voice Activity Detectorβ9,901Jul 16, 2026Updated 3 weeks ago
- Distribute and run LLMs with a single file.β25,524Aug 3, 2026Updated last week
- SoTA open-source TTSβ25,923Jul 21, 2026Updated 2 weeks ago
- Open-source desktop app for local LLMs. Text, vision, tool-calling, OpenAI/Anthropic-compatible API. 100% private.β47,543Jun 2, 2026Updated 2 months ago
- AllTalk is based on the Coqui TTS engine, similar to the Coqui_tts extension for Text generation webUI, however supports a variety of advβ¦β2,424Jan 9, 2026Updated 7 months ago
- Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Nodeβ15,032Updated this week
- Towards Human-Sounding Speechβ6,283Dec 5, 2025Updated 8 months ago
- Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.β178,129Updated this week
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Suno AI's Bark model in C/C++ for fast text-to-speech generationβ866Nov 16, 2024Updated last year
- Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model w/multiplatform CPU, AMD, NVIDIA GPU PyTorch support; voice-mixing, auto-sβ¦β5,305Updated this week
- A fast local neural text to speech engine for Mycroftβ1,263Mar 25, 2025Updated last year
- A TTS model capable of generating ultra-realistic dialogue in one pass.β19,367Nov 19, 2025Updated 8 months ago
- Local voice recording for creating Piper datasetsβ217Feb 20, 2026Updated 5 months ago
- An open source voice assistant toolkit for many human languagesβ382Dec 26, 2023Updated 2 years ago
- Silero Models: pre-trained text-to-speech models made embarrassingly simpleβ6,056Jul 31, 2026Updated last week