An end-to-end library for training audio wake-word models and deploying them in the browser.
☆45Jul 25, 2025Updated last year
Alternatives and similar repositories for hey-buddy
Users that are interested in hey-buddy are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An open source real-time AI inference engine for seamless scaling☆24Jul 2, 2025Updated last year
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- A systems programming language where automatic differentiation is a compiler pass and model parameters are explicit, growable memory.☆27Jan 5, 2026Updated 8 months ago
- SynTTS-Commands is a large-scale, multilingual (English & Chinese) synthetic speech command dataset designed for low-power Keyword Spotti…☆18Feb 5, 2026Updated 7 months ago
- ru-normalizr — лучший нормализатор русского текста без LLM. Приводит числа, даты, время, сокращения, римские цифры, символы и латиницу в …☆23Aug 7, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Arabic Grapheme-to-Phoneme (G2P) Conversion☆15Mar 15, 2025Updated last year
- Python library for evaluating and comparing generative models. Unified interface for computing quality metrics across images, videos, te…☆18Dec 18, 2025Updated 9 months ago
- Code for CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment, Interspeech 2026.☆21Aug 14, 2026Updated last month
- A universal phone recognizer that can transcribe speech in 70+ languages into IPA☆41Jun 9, 2026Updated 3 months ago
- Speech recognition module for Python, supporting several engines and APIs, online and offline.☆13Mar 9, 2022Updated 4 years ago
- Offline audiobook reader for EPUB, PDF, and TXT with local OmniVoice TTS, character voices, and synced highlighting.☆21Jul 21, 2026Updated 2 months ago
- Java Bindings for the C++ library DeepSpeech☆10Jun 4, 2020Updated 6 years ago
- ☆13May 1, 2026Updated 4 months ago
- ☆13Oct 27, 2021Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- This is a repository for a paper accepted at the 2022 IEEE Spoken Language Technology Workshop (SLT 2022)☆17Dec 1, 2022Updated 3 years ago
- ☆14Aug 19, 2024Updated 2 years ago
- Wanxiang-open-web是运行在浏览器上的三维场景设计工具,支持设计和展示两种状态。 开发语言:JavaScript。 本项目代码为三维场景设计的核心代码,完整项目(包含后端和前端)请关注wanxiang-open-service☆12May 27, 2025Updated last year
- Building actual open source including dataset Multilingual TTS more than 150 languages with Voice Cloning.☆57Updated this week
- Launch your speech synthesis within one minute.☆12May 6, 2024Updated 2 years ago
- FS Browser side Javascript module (server-to-client adaptation)☆14Mar 11, 2026Updated 6 months ago
- StyleTTS2 + Vocos as a Decoder☆13Jul 31, 2026Updated last month
- ☆22Jun 30, 2021Updated 5 years ago
- Assistance component base for Dicio assistant components☆14Apr 23, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Keyword Spotting using BCResNet and Arcface Loss☆13Jan 28, 2022Updated 4 years ago
- This will hold the crowdsourcing platform to be used to store voice data from various speakers which will act as input dataset for speech…☆17Mar 6, 2023Updated 3 years ago
- Using OpenVINO to speed up MeloTTS inference☆15Nov 1, 2024Updated last year
- ☆18Apr 28, 2021Updated 5 years ago
- wake word spotting with kaldi☆19Dec 3, 2020Updated 5 years ago
- The dataset construction pipeline for WordVoice-5A☆19Jul 17, 2026Updated 2 months ago
- Parallelized automatic corpus collection for ASR. Forked from https://github.com/EgorLakomkin/KTSpeechCrawler☆23Mar 21, 2021Updated 5 years ago
- Open-source AI that watches and hears your screen and reacts live as any personality you design — for faceless content, autonomous AI str…☆38Jul 7, 2026Updated 2 months ago
- mnn tts demo.☆20May 7, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official Implementation for "Age-Dependent Face Diversification via Latent Space Analysis" (CGI2023)☆15Jan 7, 2025Updated last year
- 🌼 Daisy-TTS: Simulating Wider Spectrum of Emotions via Prosody Embedding Decomposition☆14Nov 15, 2025Updated 10 months ago
- This is a demo for SOTA vocal separation models. Upload an audio file and the model will separate the vocals from the background music. …☆18Jul 25, 2024Updated 2 years ago
- Finally, some decent sample sentences☆24Dec 3, 2023Updated 2 years ago
- Ubuntu Root Filesystem creation tools☆11Nov 4, 2025Updated 10 months ago
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IW…☆18Nov 30, 2022Updated 3 years ago
- Wake word detection with custom phrases without model training☆60Mar 8, 2026Updated 6 months ago