SynTTS-Commands is a large-scale, multilingual (English & Chinese) synthetic speech command dataset designed for low-power Keyword Spotting (KWS) tasks. Generated using state-of-the-art TTS technology (CosyVoice 2), it addresses the data scarcity bottleneck in TinyML and Edge AI.
☆17Feb 5, 2026Updated 6 months ago
Alternatives and similar repositories for SynTTS-Commands-Official
Users that are interested in SynTTS-Commands-Official are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [Tiny KWS] SparkNet: Sparse Binarization for Fast Keyword Spotting☆20Aug 26, 2025Updated 11 months ago
- Test Framework for few-shot open set KWS☆45Nov 8, 2024Updated last year
- A universal phone recognizer that can transcribe speech in 70+ languages into IPA☆31Jun 9, 2026Updated 2 months ago
- this repository contains a Colab notebook to classify the heart sound as normal or abnormal☆11Jul 5, 2020Updated 6 years ago
- Code for the Interspeech 2024 paper "MM-KWS: Multi-modal Prompts for Multilingual User-defined Keyword Spotting"☆51Jan 24, 2026Updated 6 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆94Jun 25, 2025Updated last year
- A lightweight, wake word detection engine. Train custom, high-accuracy models with minimal effort.☆114Jun 15, 2026Updated last month
- Collection of PyTorch implementations of Spoken Keyword Spotting presented in research papers.☆42Apr 5, 2024Updated 2 years ago
- Wake word detection with custom phrases without model training☆56Mar 8, 2026Updated 5 months ago
- Tiny Transducer: A Highly-Efficient Speech Recognition Model on Edge Devices☆30Aug 4, 2022Updated 4 years ago
- Code for CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment, Interspeech 2026.☆20Jun 9, 2026Updated 2 months ago
- ☆100May 31, 2023Updated 3 years ago
- The new dataset, which has been published in Scientific Data (Nature Portfolio), is divided into a training set (train: 882), a validatio…☆17May 1, 2026Updated 3 months ago
- ru-normalizr — лучший нормализатор русского текста без LLM. Приводит числа, даты, время, сокращения, римские цифры, символы и латиницу в …☆19Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆25Mar 8, 2026Updated 5 months ago
- Arabic Grapheme-to-Phoneme (G2P) Conversion☆15Mar 15, 2025Updated last year
- FNSE-SBGAN: Far-field Speech Enhancement with Schrödinger Bridge and Generative Adversarial Networks☆20May 12, 2025Updated last year
- Voice Framework☆18Jan 21, 2026Updated 6 months ago
- Official Repository for "Global Rotation Equivariant Phase Modeling for Speech Enhancement with Deep Magnitude-Phase Interaction"☆20Jun 25, 2026Updated last month
- A real-time voice conversion model based on VITS.☆16Aug 1, 2024Updated 2 years ago
- An open-source, easily accessible package for training and deploying Speech-to-Intent models on microcontrollers and SBCs☆51Mar 14, 2024Updated 2 years ago
- KWS demo based on CTC prefix beam search.☆19Oct 21, 2023Updated 2 years ago
- [Tiny VAD] SG-VAD: Stochastic Gates Based Speech Activity Detection☆40Mar 24, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- SALT: STANDARDIZED AUDIO EVENT LABEL TAXONOMY☆16Nov 28, 2024Updated last year
- All-in-one Speech Transcription☆11Jun 5, 2026Updated 2 months ago
- speex aec kalman filter☆15Mar 17, 2024Updated 2 years ago
- The code about “LABNet: A Lightweight Attentive Beamforming Network for Ad-hoc Multichannel Microphone Invariant Real-Time Speech Enhance…☆49Oct 10, 2025Updated 10 months ago
- ☆12Jun 17, 2017Updated 9 years ago
- LLaSE: Maximizing Acoustic Preservation for LLaMA based Speech Enhancement☆16Jul 11, 2025Updated last year
- ☆24Jul 29, 2024Updated 2 years ago
- TinyML Papers☆15Mar 11, 2025Updated last year
- Building actual open source including dataset Multilingual TTS more than 150 languages with Voice Cloning.☆56Jul 14, 2026Updated 3 weeks ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- An upgrade framework for train and validate compare with icefall using Lightning.☆16Mar 26, 2025Updated last year
- ☆27Aug 29, 2025Updated 11 months ago
- Keyword Spotting using BCResNet and Arcface Loss☆13Jan 28, 2022Updated 4 years ago
- ☆11May 5, 2025Updated last year
- BC-ResNet for Keyword Spotting☆44Jan 11, 2022Updated 4 years ago
- ☆32Aug 1, 2021Updated 5 years ago
- Label smoothed Aggregation cross entropy loss for generalisation in sequence to sequence tasks.☆14Dec 17, 2019Updated 6 years ago