This repository contains the Code for SOTA model on Google Speech Command V2 dataset.
☆15Sep 28, 2023Updated 2 years ago
Alternatives and similar repositories for GoogleSpeechCommandLowFootprint
Users that are interested in GoogleSpeechCommandLowFootprint are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- OLaPh (Optimal Language Phonemizer) is a multilingual phonemization framework that converts text into phonemes surpassing the quality of …☆22Jul 20, 2026Updated 3 weeks ago
- ☆100May 31, 2023Updated 3 years ago
- A time delay estimation method for event-based time-series data. Time delay estimation is also known as the correction of time offsets an…☆16Dec 3, 2025Updated 8 months ago
- ☆25Jul 19, 2026Updated 3 weeks ago
- [INTERSPEECH 2026] Official code for "Balalaika: Data-Centric, Prosody-Aware Annotation Pipeline for Russian Speech"☆21Jul 19, 2026Updated 3 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official implementation of "PhonMatchNet: Phoneme-Guided Zero-Shot Keyword Spotting for User-Defined Keywords" (INTERSPEECH 2023)☆63Jun 3, 2024Updated 2 years ago
- offical code for Dense-TSNet☆12Sep 17, 2024Updated last year
- ☆90May 27, 2023Updated 3 years ago
- golang vad (voice activity detection) library based on webrtc☆12Dec 13, 2021Updated 4 years ago
- Neural network auditory processing code in Go focused on filtering speech wav files via mel filters☆11Jan 22, 2024Updated 2 years ago
- Official repository for the paper "MambAttention: Mamba with Multi-Head Attention for Generalizable Single-Channel Speech Enhancement" (A…☆35Mar 25, 2026Updated 4 months ago
- ☆11Oct 20, 2022Updated 3 years ago
- ☆14Jun 19, 2019Updated 7 years ago
- Noise Reduction Preprocessing-based Fully Automatic Diagonal Loading Method for Robust Adaptive Beamforming☆13Feb 24, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A lightweight, efficient variation of the StyleTTS 2 text‐to‐speech model.☆50May 22, 2025Updated last year
- This repo contains required files for the INTERSPEECH 2022 Audio Deep Packet Loss Concealment (PLC) Challenge.☆92Feb 13, 2026Updated 5 months ago
- ☆11May 5, 2025Updated last year
- BC-ResNet for Keyword Spotting☆44Jan 11, 2022Updated 4 years ago
- Inframon - Local Network Server Monitor for macOS and Linux AMD Machines!☆17Nov 13, 2024Updated last year
- A STFT/iSTFT written up in PyTorch using 1D Convolutions☆32Jul 9, 2024Updated 2 years ago
- ☆18Dec 27, 2023Updated 2 years ago
- Slick video review note taking app 🎬☆12Jul 26, 2026Updated 2 weeks ago
- [INTERSPEECH 2025 Oral]Official code for "Accelerating Diffusion-based Text-to-Speech Model Training with Dual Modality Alignment"☆67Jun 16, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The official implementation of the method discussed in the paper Improving Spoken Language Identification with Map-Mix(work accepted at I…☆18Feb 17, 2023Updated 3 years ago
- Go port of the metaphone3 algorithm☆25Sep 3, 2019Updated 6 years ago
- ☆15Jan 18, 2021Updated 5 years ago
- Build tensorflow lite static library or shared library by android ndk☆13Sep 14, 2018Updated 7 years ago
- ArcFace 3.0的Android Demo☆12Dec 16, 2019Updated 6 years ago
- A converter from Arpabet to IPA (see https://en.wikipedia.org/wiki/Arpabet)☆17Jan 2, 2018Updated 8 years ago
- LibCP -- A Library for Conformal Prediction☆13Feb 26, 2015Updated 11 years ago
- Script to simulate room impulse responses☆16Sep 29, 2016Updated 9 years ago
- Token-Level Ensemble Distillation for Grapheme-to-Phoneme Conversion☆20Jul 9, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation of Generalized Cross Correlation with Phase Transform (GCC-PHAT) library in C/C++.☆20Jul 8, 2019Updated 7 years ago
- Rank 7th/1817 in the 2018 iFLYTEK AI Developer Challenge with acc 0.82 for the ten Chinese dialects classification task, this code was p…☆14Nov 19, 2023Updated 2 years ago
- Code for "Error-driven Fixed-Budget ASR Personalization for Accented Speakers" in ICASSP 2021☆11Jun 13, 2021Updated 5 years ago
- [TMLR] Official PyTorch implementation of paper "Quantization Variation: A New Perspective on Training Transformers with Low-Bit Precisio…☆50Sep 27, 2024Updated last year
- A Persian Word2Vec Model trained by Wikipedia articles☆10Jan 5, 2018Updated 8 years ago
- FastAudio is a Learnable Audio Frontend team Magnum's designed for the ASVspoof 2021 challenge☆46May 6, 2023Updated 3 years ago
- SynTTS-Commands is a large-scale, multilingual (English & Chinese) synthetic speech command dataset designed for low-power Keyword Spotti…☆17Feb 5, 2026Updated 6 months ago