Code for the Interspeech 2024 paper "MM-KWS: Multi-modal Prompts for Multilingual User-defined Keyword Spotting"
☆53Jan 24, 2026Updated 7 months ago
Alternatives and similar repositories for MM-KWS
Users that are interested in MM-KWS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 开源自定义唤醒词☆17Dec 24, 2025Updated 8 months ago
- Official implementation of "PhonMatchNet: Phoneme-Guided Zero-Shot Keyword Spotting for User-Defined Keywords" (INTERSPEECH 2023)☆63Jun 3, 2024Updated 2 years ago
- Official code for Metric learning for user-defined keyword spotting☆40Feb 21, 2024Updated 2 years ago
- Recipe for LibriPhrase☆40Sep 2, 2023Updated 3 years ago
- Test Framework for few-shot open set KWS☆45Nov 8, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Test-time adaptation for speech recognition model by single utterance. The official implementation of "Listen, Adapt, Better WER: Source-…☆23Apr 1, 2022Updated 4 years ago
- ☆95Jun 25, 2025Updated last year
- ☆33Aug 10, 2022Updated 4 years ago
- E2E ASR system☆14Oct 20, 2022Updated 3 years ago
- SynTTS-Commands is a large-scale, multilingual (English & Chinese) synthetic speech command dataset designed for low-power Keyword Spotti…☆18Feb 5, 2026Updated 7 months ago
- ☆27Aug 29, 2025Updated last year
- Collection of PyTorch implementations of Spoken Keyword Spotting presented in research papers.☆45Apr 5, 2024Updated 2 years ago
- Generate synthetic wind noise signals based on a wind speed profile (Python)☆55Apr 23, 2024Updated 2 years ago
- End-to-End Speech Processing Toolkit☆16Jan 20, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Keyword spotting and forced alignment in any language☆102Jun 15, 2026Updated 2 months ago
- 基于两阶段的声学回声消除系统 A Two-Stage-Based Acoustic Echo Cancellation System☆17Feb 22, 2026Updated 6 months ago
- The code about “LABNet: A Lightweight Attentive Beamforming Network for Ad-hoc Multichannel Microphone Invariant Real-Time Speech Enhance…☆52Updated this week
- Production First and Production Ready End-to-End Keyword Spotting Toolkit☆761Jul 23, 2026Updated last month
- ☆163May 26, 2023Updated 3 years ago
- ☆16Feb 22, 2025Updated last year
- Pytorch implementation of BiFSMNv2, TNNLS 2023☆37Feb 10, 2023Updated 3 years ago
- PyTorch reimplementation of "Keyword Transformer: A Self-Attention Model for Keyword Spotting"☆16Jul 23, 2021Updated 5 years ago
- Few-shot Keyword Spotting in Any Language and Multilingual Spoken Word Corpus☆190Dec 6, 2024Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Official implementation of Efficient Speech Separation Framework Based on Neural State-Space Models☆29Feb 25, 2026Updated 6 months ago
- ☆102May 31, 2023Updated 3 years ago
- This repository contains code for applying Data2Vec to pretrain Keyword Transformer model as described in "Improving Label-Deficient Keyw…☆32Mar 6, 2025Updated last year
- Official Repository for "Global Rotation Equivariant Phase Modeling for Speech Enhancement with Deep Magnitude-Phase Interaction"☆20Jun 25, 2026Updated 2 months ago
- [Tiny KWS] SparkNet: Sparse Binarization for Fast Keyword Spotting☆20Aug 26, 2025Updated last year
- This repository is a curated list of awesome Speech Keyword Spotting (Wake-Up Word Detection).☆294May 23, 2022Updated 4 years ago
- Official repository of Spiking-FullSubNet, the Intel N-DNS Challenge Algorithmic Track Winner.☆144Jan 28, 2026Updated 7 months ago
- ☆20Jun 16, 2026Updated 2 months ago
- [ICLR 2026] StableToken: A state-of-the-art noise-robust semantic speech tokenizer featuring Voting-LFQ for resilient SpeechLLMs.☆34Feb 27, 2026Updated 6 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Speech-MASSIVE is a multilingual Spoken Language Understanding (SLU) dataset comprising the speech counterpart for a portion of the MASSI…☆25Oct 8, 2025Updated 10 months ago
- ☆29Jul 25, 2026Updated last month
- LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement☆106Apr 1, 2025Updated last year
- Perceptual Contrast Stretching on Target Feature for Speech Enhancement (Accepted by INTERSPEECH 2022)☆73May 11, 2024Updated 2 years ago
- A list of datasets made available by members of the Aalto Acoustics Lab☆32Updated this week
- Toolbox for Evaluation of AEC/AES Systems☆43Feb 18, 2026Updated 6 months ago
- Few-Shot Keyword Spotting☆73Apr 11, 2021Updated 5 years ago