Babylon.cpp is a C and C++ library for grapheme to phoneme conversion and text to speech synthesis. For phonemization a ONNX runtime port of the DeepPhonemizer model is used. For speech synthesis VITS models are used. Piper models are compatible after a conversion script is run.
β41Apr 14, 2026Updated 4 months ago
Alternatives and similar repositories for babylon
Users that are interested in babylon are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π± Flutter demo app for Arabic TTS ποΈ β ONNX-based offline speech synthesis πβ17May 3, 2025Updated last year
- ποΈ Arabic TTS models (FastPitch, Mixer-TTS) in the ONNX format β Python package for offline speech synthesis ππ¦β45Jun 20, 2026Updated 2 months ago
- Using OpenVINO to speed up MeloTTS inferenceβ15Nov 1, 2024Updated last year
- mnn tts demo.β19May 7, 2025Updated last year
- Colab notebooks for Next-gen Kaldiβ31Oct 12, 2025Updated 10 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An espeak-compatible, permissively-licensed IPA phonemizer (G2P) based on DeepPhonemizer. Usable as a drop-in replacement for espeak's GPβ¦β112Mar 15, 2026Updated 5 months ago
- Speech recognition module for Python, supporting several engines and APIs, online and offline.β13Mar 9, 2022Updated 4 years ago
- Java Bindings for the C++ library DeepSpeechβ10Jun 4, 2020Updated 6 years ago
- β13May 1, 2026Updated 3 months ago
- This is a balanced dataset for English homograph disambiguation (HD), generated with Meta's Llama 2-Chat 70B model.β22Jan 22, 2024Updated 2 years ago
- β33Nov 27, 2021Updated 4 years ago
- ESLTTS datasetβ16Feb 6, 2025Updated last year
- β18Nov 19, 2025Updated 9 months ago
- Launch your speech synthesis within one minute.β12May 6, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- β33Aug 6, 2021Updated 5 years ago
- Add Arabic diacritics (tashkeel/harakat) using Rust/Python/C++/WASM and NLP modelsβ52Oct 4, 2025Updated 10 months ago
- Uses the excellent silero VAD with onnxruntime C api for fast detection of audio segments with speechβ17Sep 20, 2024Updated last year
- β22Jun 30, 2021Updated 5 years ago
- Free, local voice-to-text for Windows & macOS. No cloud, no account, no subscription.β15Jul 11, 2026Updated last month
- Assistance component base for Dicio assistant componentsβ14Apr 23, 2026Updated 4 months ago
- VITS Inference using ONNX Runtime on C++β13Dec 25, 2023Updated 2 years ago
- This will hold the crowdsourcing platform to be used to store voice data from various speakers which will act as input dataset for speechβ¦β17Mar 6, 2023Updated 3 years ago
- zero-shot realtime TTS system, fully offline, free and open sourceβ56Apr 18, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- π§ Bio-Agent OS: π»π³ Bio-Inspired Memory Framework for AI Agents (OpenClaw/ERP). Researched & Developed by Dev Tuan Anh Ha (Locaith Soluβ¦β20Aug 20, 2026Updated last week
- A Benchmark Corpus for Low-Resource Cantonese Punctuation Restoration from Speech Transcriptsβ15Dec 3, 2024Updated last year
- β18Apr 28, 2021Updated 5 years ago
- wake word spotting with kaldiβ19Dec 3, 2020Updated 5 years ago
- A framework for creating voice based agents. Integrations LLMs with speech recognition and text-to-speechβ35May 1, 2024Updated 2 years ago
- β40Aug 15, 2021Updated 5 years ago
- Lightweight on-device keyword spotting engine for iOS using CoreML and real-time audio streaming.β16Jun 14, 2025Updated last year
- Parallelized automatic corpus collection for ASR. Forked from https://github.com/EgorLakomkin/KTSpeechCrawlerβ23Mar 21, 2021Updated 5 years ago
- β14Aug 19, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Finally, some decent sample sentencesβ24Dec 3, 2023Updated 2 years ago
- Code and Resources for "LLM-Powered Grapheme-to-Phoneme Conversion: Benchmark and Case Study", introducing methods to leverage LLMs for Gβ¦β19May 21, 2025Updated last year
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IWβ¦β18Nov 30, 2022Updated 3 years ago
- Clean and modernized implementation of FastSpeech2/LightSpeech using IPAβ20Aug 16, 2024Updated 2 years ago
- β22Sep 24, 2018Updated 7 years ago
- Unofficial implementation of ConvNeXt-TTS powered by lightningβ18Oct 20, 2024Updated last year
- StyleTTS2 + Vocos as a Decoderβ13Jul 31, 2026Updated last month