☆94Jun 25, 2025Updated last year
Alternatives and similar repositories for KWStreamingSearch
Users that are interested in KWStreamingSearch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for the Interspeech 2024 paper "MM-KWS: Multi-modal Prompts for Multilingual User-defined Keyword Spotting"☆51Jan 24, 2026Updated 6 months ago
- Materials of public talks given By SJTU X-LANCE members☆14Dec 3, 2022Updated 3 years ago
- ☆15Sep 16, 2024Updated last year
- 开源自定义唤醒词☆17Dec 24, 2025Updated 7 months ago
- TASU: A New Style of Alignment of Speech LLM with only Text Training Data, zero-shot on ASR and Other SU tasks☆27Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- SynTTS-Commands is a large-scale, multilingual (English & Chinese) synthetic speech command dataset designed for low-power Keyword Spotti…☆17Feb 5, 2026Updated 5 months ago
- Target Speaker Extraction Toolkit☆299Oct 4, 2025Updated 9 months ago
- Exploring Binary Classification Loss for Speaker Verification☆18Jul 18, 2023Updated 3 years ago
- A toolkit for speaker diarization.☆504May 29, 2026Updated last month
- KWS demo based on CTC prefix beam search.☆19Oct 21, 2023Updated 2 years ago
- Collection of PyTorch implementations of Spoken Keyword Spotting presented in research papers.☆41Apr 5, 2024Updated 2 years ago
- Official implementation of "PhonMatchNet: Phoneme-Guided Zero-Shot Keyword Spotting for User-Defined Keywords" (INTERSPEECH 2023)☆63Jun 3, 2024Updated 2 years ago
- LLaSE: Maximizing Acoustic Preservation for LLaMA based Speech Enhancement☆16Jul 11, 2025Updated last year
- [ACL 2025 Main] UniCodec: a unified audio codec with a single codebook to support multi-domain audio data, including speech, music, and s…☆157May 30, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- We Speech Transcript based on LLM, in 300 lines of code.☆182Jun 20, 2025Updated last year
- [AAAI 2024] Code for CTX-vec2wav in UniCATS☆130Jun 11, 2024Updated 2 years ago
- ☆26Aug 29, 2025Updated 10 months ago
- Code for vec2wav 2.0, a speech token vocoder for VC. Paper: https://arxiv.org/abs/2409.01995☆79Dec 3, 2024Updated last year
- An N-gram punctuator for Chinese and English.☆18Oct 14, 2025Updated 9 months ago
- 5Hz Deep-Compression Speech VAE for AR-Diffusion and CALMs☆57Nov 19, 2025Updated 8 months ago
- ☆10Mar 22, 2023Updated 3 years ago
- Production first, nn-based on-device signal processing toolkit.☆63May 30, 2023Updated 3 years ago
- Single Channel Speech Enhancement Methods and Toolbox☆55Jun 28, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [Tiny KWS] SparkNet: Sparse Binarization for Fast Keyword Spotting☆20Aug 26, 2025Updated 10 months ago
- LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement☆105Apr 1, 2025Updated last year
- Multi-Stage Face-Voice Association Learning with Keynote Speaker Diarization (ACM MM 2024)☆22Jul 25, 2024Updated 2 years ago
- Paper Claw sends personalized daily research digests from arXiv and beyond straight to your inbox, featuring customizable categories, int…☆32Updated this week
- FINALLY: Fast and universal speech enhancement model delivering studio-quality audio for a wide range of recordings.☆28Apr 1, 2026Updated 3 months ago
- HoliTok:A Coutinuous Holistic Tokenization with Robust Dual Capabilities of Speech Generation and Understanding☆39Jun 8, 2026Updated last month
- end-to-end text to audio scene generation model☆50Jun 16, 2026Updated last month
- ☆15Jul 23, 2024Updated 2 years ago
- BAE-NET: A LOW COMPLEXITY AND HIGH FIDELITY BANDWIDTH-ADAPTIVE NEURAL NETWORK FOR SPEECH SUPER-RESOLUTION☆80Aug 20, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A CSRankings-like index for speech researchers☆35Oct 16, 2024Updated last year
- A Framework for Speech, Language, Audio, Music Processing with Large Language Model☆1,049Jan 15, 2026Updated 6 months ago
- Supporting code for the paper "A study on more realistic room simulation for far-field keyword spotting".☆34Oct 27, 2020Updated 5 years ago
- This is the official implementation of the LiSenNet☆162Nov 15, 2024Updated last year
- The implementation of TaylorBeamformer, which is in submission to Interspeech2022☆49Jun 10, 2022Updated 4 years ago
- ☆39Jul 4, 2024Updated 2 years ago
- Official repository of Spiking-FullSubNet, the Intel N-DNS Challenge Algorithmic Track Winner.☆142Jan 28, 2026Updated 5 months ago