Babylon.cpp is a C and C++ library for grapheme to phoneme conversion and text to speech synthesis. For phonemization a ONNX runtime port of the DeepPhonemizer model is used. For speech synthesis VITS models are used. Piper models are compatible after a conversion script is run.
β39Apr 14, 2026Updated 3 months ago
Alternatives and similar repositories for babylon
Users that are interested in babylon are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π± Flutter demo app for Arabic TTS ποΈ β ONNX-based offline speech synthesis πβ17May 3, 2025Updated last year
- ποΈ Arabic TTS models (FastPitch, Mixer-TTS) in the ONNX format β Python package for offline speech synthesis ππ¦β44Jun 20, 2026Updated last month
- Using OpenVINO to speed up MeloTTS inferenceβ15Nov 1, 2024Updated last year
- mnn tts demo.β19May 7, 2025Updated last year
- An espeak-compatible, permissively-licensed IPA phonemizer (G2P) based on DeepPhonemizer. Usable as a drop-in replacement for espeak's GPβ¦β111Mar 15, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Colab notebooks for Next-gen Kaldiβ31Oct 12, 2025Updated 9 months ago
- Speech recognition module for Python, supporting several engines and APIs, online and offline.β13Mar 9, 2022Updated 4 years ago
- Java Bindings for the C++ library DeepSpeechβ10Jun 4, 2020Updated 6 years ago
- β13May 1, 2026Updated 2 months ago
- This is a balanced dataset for English homograph disambiguation (HD), generated with Meta's Llama 2-Chat 70B model.β22Jan 22, 2024Updated 2 years ago
- β33Nov 27, 2021Updated 4 years ago
- β13Oct 27, 2021Updated 4 years ago
- ESLTTS datasetβ16Feb 6, 2025Updated last year
- β18Nov 19, 2025Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Launch your speech synthesis within one minute.β12May 6, 2024Updated 2 years ago
- β33Aug 6, 2021Updated 4 years ago
- Add Arabic diacritics (tashkeel/harakat) using Rust/Python/C++/WASM and NLP modelsβ50Oct 4, 2025Updated 9 months ago
- Uses the excellent silero VAD with onnxruntime C api for fast detection of audio segments with speechβ16Sep 20, 2024Updated last year
- β22Jun 30, 2021Updated 5 years ago
- Free, local voice-to-text for Windows & macOS. No cloud, no account, no subscription.β15Jul 11, 2026Updated last week
- Assistance component base for Dicio assistant componentsβ13Apr 23, 2026Updated 2 months ago
- VITS Inference using ONNX Runtime on C++β13Dec 25, 2023Updated 2 years ago
- This will hold the crowdsourcing platform to be used to store voice data from various speakers which will act as input dataset for speechβ¦β17Mar 6, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- zero-shot realtime TTS system, fully offline, free and open sourceβ55Apr 18, 2025Updated last year
- π§ Bio-Agent OS: π»π³ Bio-Inspired Memory Framework for AI Agents (OpenClaw/ERP). Researched & Developed by Dev Tuan Anh Ha (Locaith Soluβ¦β20Apr 21, 2026Updated 3 months ago
- A Benchmark Corpus for Low-Resource Cantonese Punctuation Restoration from Speech Transcriptsβ15Dec 3, 2024Updated last year
- β18Apr 28, 2021Updated 5 years ago
- wake word spotting with kaldiβ19Dec 3, 2020Updated 5 years ago
- A framework for creating voice based agents. Integrations LLMs with speech recognition and text-to-speechβ35May 1, 2024Updated 2 years ago
- β40Aug 15, 2021Updated 4 years ago
- CodecHub: A Unified Library for Codec Modelsβ25Dec 24, 2025Updated 6 months ago
- Lightweight on-device keyword spotting engine for iOS using CoreML and real-time audio streaming.β16Jun 14, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Parallelized automatic corpus collection for ASR. Forked from https://github.com/EgorLakomkin/KTSpeechCrawlerβ23Mar 21, 2021Updated 5 years ago
- β14Aug 19, 2024Updated last year
- Finally, some decent sample sentencesβ24Dec 3, 2023Updated 2 years ago
- Code and Resources for "LLM-Powered Grapheme-to-Phoneme Conversion: Benchmark and Case Study", introducing methods to leverage LLMs for Gβ¦β19May 21, 2025Updated last year
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IWβ¦β18Nov 30, 2022Updated 3 years ago
- Clean and modernized implementation of FastSpeech2/LightSpeech using IPAβ18Aug 16, 2024Updated last year
- β22Sep 24, 2018Updated 7 years ago