An end-to-end library for training audio wake-word models and deploying them in the browser.
☆45Jul 25, 2025Updated last year
Alternatives and similar repositories for hey-buddy
Users that are interested in hey-buddy are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An open source real-time AI inference engine for seamless scaling☆23Jul 2, 2025Updated last year
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- Code for CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment, Interspeech 2026.☆20Jun 9, 2026Updated 2 months ago
- SynTTS-Commands is a large-scale, multilingual (English & Chinese) synthetic speech command dataset designed for low-power Keyword Spotti…☆17Feb 5, 2026Updated 6 months ago
- ru-normalizr — лучший нормализатор русского текста без LLM. Приводит числа, даты, время, сокращения, римские цифры, символы и латиницу в …☆19Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Arabic Grapheme-to-Phoneme (G2P) Conversion☆15Mar 15, 2025Updated last year
- Python library for evaluating and comparing generative models. Unified interface for computing quality metrics across images, videos, te…☆17Dec 18, 2025Updated 7 months ago
- A universal phone recognizer that can transcribe speech in 70+ languages into IPA☆31Jun 9, 2026Updated 2 months ago
- Speech recognition module for Python, supporting several engines and APIs, online and offline.☆13Mar 9, 2022Updated 4 years ago
- Voice mixer and modifier for SuperTonic TTS☆33Nov 25, 2025Updated 8 months ago
- Offline audiobook reader for EPUB, PDF, and TXT with local OmniVoice TTS, character voices, and synced highlighting.☆20Jul 21, 2026Updated 3 weeks ago
- Java Bindings for the C++ library DeepSpeech☆10Jun 4, 2020Updated 6 years ago
- ☆13May 1, 2026Updated 3 months ago
- ☆13Oct 27, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is a repository for a paper accepted at the 2022 IEEE Spoken Language Technology Workshop (SLT 2022)☆17Dec 1, 2022Updated 3 years ago
- ☆14Aug 19, 2024Updated last year
- Building actual open source including dataset Multilingual TTS more than 150 languages with Voice Cloning.☆56Jul 14, 2026Updated 3 weeks ago
- Launch your speech synthesis within one minute.☆12May 6, 2024Updated 2 years ago
- StyleTTS2 + Vocos as a Decoder☆13Jul 31, 2026Updated last week
- ☆22Jun 30, 2021Updated 5 years ago
- Assistance component base for Dicio assistant components☆14Apr 23, 2026Updated 3 months ago
- Keyword Spotting using BCResNet and Arcface Loss☆13Jan 28, 2022Updated 4 years ago
- This will hold the crowdsourcing platform to be used to store voice data from various speakers which will act as input dataset for speech…☆17Mar 6, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Using OpenVINO to speed up MeloTTS inference☆15Nov 1, 2024Updated last year
- wake word spotting with kaldi☆19Dec 3, 2020Updated 5 years ago
- mnn tts demo.☆19May 7, 2025Updated last year
- The dataset construction pipeline for WordVoice-5A☆19Jul 17, 2026Updated 3 weeks ago
- Lightweight on-device keyword spotting engine for iOS using CoreML and real-time audio streaming.☆16Jun 14, 2025Updated last year
- Open-source AI that watches and hears your screen and reacts live as any personality you design — for faceless content, autonomous AI str…☆27Jul 7, 2026Updated last month
- 🌼 Daisy-TTS: Simulating Wider Spectrum of Emotions via Prosody Embedding Decomposition☆14Nov 15, 2025Updated 8 months ago
- This is a demo for SOTA vocal separation models. Upload an audio file and the model will separate the vocals from the background music. …☆18Jul 25, 2024Updated 2 years ago
- Finally, some decent sample sentences☆24Dec 3, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IW…☆18Nov 30, 2022Updated 3 years ago
- Wake word detection with custom phrases without model training☆56Mar 8, 2026Updated 5 months ago
- Implementation of RIFT-SVC, a singing voice conversion model based on Rectified Flow Transformer.☆70Nov 10, 2025Updated 9 months ago
- A pitch detection model trained to be robust against noise and reverberation environments.☆27Jan 21, 2025Updated last year
- ☆22Sep 24, 2018Updated 7 years ago
- Lite Voice Terminal, an "offline smart speaker" solution powered by on-premise ASR server (vosk API / kaldi engine)☆19Feb 29, 2024Updated 2 years ago
- a Neural Vocoder supporting Ring Attention, Conformer and NSF.☆25Aug 1, 2025Updated last year