An end-to-end library for training audio wake-word models and deploying them in the browser.
☆44Jul 25, 2025Updated 11 months ago
Alternatives and similar repositories for hey-buddy
Users that are interested in hey-buddy are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An open source real-time AI inference engine for seamless scaling☆23Jul 2, 2025Updated last year
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- Code for CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment, Interspeech 2026.☆20Jun 9, 2026Updated last month
- SynTTS-Commands is a large-scale, multilingual (English & Chinese) synthetic speech command dataset designed for low-power Keyword Spotti…☆17Feb 5, 2026Updated 5 months ago
- ru-normalizr — лучший нормализатор русского текста без LLM. Приводит числа, даты, время, сокращения, римские цифры, символы и латиницу в …☆17Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Arabic Grapheme-to-Phoneme (G2P) Conversion☆16Mar 15, 2025Updated last year
- A universal phone recognizer that can transcribe speech in 70+ languages into IPA☆21Jun 9, 2026Updated last month
- Offline audiobook reader for EPUB, PDF, and TXT with local OmniVoice TTS, character voices, and synced highlighting.☆17May 12, 2026Updated 2 months ago
- Voice mixer and modifier for SuperTonic TTS☆32Nov 25, 2025Updated 7 months ago
- Speech recognition module for Python, supporting several engines and APIs, online and offline.☆13Mar 9, 2022Updated 4 years ago
- Java Bindings for the C++ library DeepSpeech☆10Jun 4, 2020Updated 6 years ago
- ☆13May 1, 2026Updated 2 months ago
- ☆13Oct 27, 2021Updated 4 years ago
- This is a repository for a paper accepted at the 2022 IEEE Spoken Language Technology Workshop (SLT 2022)☆17Dec 1, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Using OpenVINO to speed up MeloTTS inference☆15Nov 1, 2024Updated last year
- ☆14Aug 19, 2024Updated last year
- Building actual open source including dataset Multilingual TTS more than 150 languages with Voice Cloning.☆54Jul 14, 2026Updated last week
- Launch your speech synthesis within one minute.☆12May 6, 2024Updated 2 years ago
- StyleTTS2 + Vocos as a Decoder☆13Mar 24, 2025Updated last year
- ☆22Jun 30, 2021Updated 5 years ago
- Assistance component base for Dicio assistant components☆13Apr 23, 2026Updated 2 months ago
- This will hold the crowdsourcing platform to be used to store voice data from various speakers which will act as input dataset for speech…☆17Mar 6, 2023Updated 3 years ago
- ☆18Apr 28, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- wake word spotting with kaldi☆19Dec 3, 2020Updated 5 years ago
- mnn tts demo.☆19May 7, 2025Updated last year
- The dataset construction pipeline for WordVoice-5A☆15Updated this week
- Open-source AI that watches and hears your screen and reacts live as any personality you design — for faceless content, autonomous AI str…☆21Jul 7, 2026Updated 2 weeks ago
- Parallelized automatic corpus collection for ASR. Forked from https://github.com/EgorLakomkin/KTSpeechCrawler☆23Mar 21, 2021Updated 5 years ago
- Official Implementation for "Age-Dependent Face Diversification via Latent Space Analysis" (CGI2023)☆15Jan 7, 2025Updated last year
- This is a demo for SOTA vocal separation models. Upload an audio file and the model will separate the vocals from the background music. …☆18Jul 25, 2024Updated last year
- 🌼 Daisy-TTS: Simulating Wider Spectrum of Emotions via Prosody Embedding Decomposition☆14Nov 15, 2025Updated 8 months ago
- Finally, some decent sample sentences☆24Dec 3, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Wake word detection with custom phrases without model training☆56Mar 8, 2026Updated 4 months ago
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IW…☆18Nov 30, 2022Updated 3 years ago
- ☆15May 13, 2026Updated 2 months ago
- A pitch detection model trained to be robust against noise and reverberation environments.☆27Jan 21, 2025Updated last year
- ☆22Sep 24, 2018Updated 7 years ago
- Lite Voice Terminal, an "offline smart speaker" solution powered by on-premise ASR server (vosk API / kaldi engine)☆19Feb 29, 2024Updated 2 years ago
- ☆16Sep 6, 2025Updated 10 months ago