Offline streaming speech-to-text in the browser
☆27Aug 28, 2025Updated 11 months ago
Alternatives and similar repositories for wasm-speech-streaming
Users that are interested in wasm-speech-streaming are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Oct 11, 2024Updated last year
- Text-to-text alignment algorithm for speech recognition error analysis.☆32Jun 23, 2026Updated last month
- ☆30Apr 29, 2026Updated 3 months ago
- Open-weights voice acting pipeline combining zero-shot voice cloning with natural-language direction. Provide a reference voice (or gener…☆17May 25, 2026Updated 2 months ago
- ☆17Dec 18, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Parallelized automatic corpus collection for ASR. Forked from https://github.com/EgorLakomkin/KTSpeechCrawler☆23Mar 21, 2021Updated 5 years ago
- Forced alignment decoder for Whisper.☆16Mar 13, 2024Updated 2 years ago
- ☆19Mar 22, 2024Updated 2 years ago
- ☆20Sep 2, 2024Updated last year
- Train no-reference speech quality estimators with multiple datasets via learned, per-dataset alignments.☆18Aug 1, 2025Updated last year
- High-performance, semantic turn detection for conversational AI☆45Oct 1, 2025Updated 10 months ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- 🎵 muse: Music Separation☆11Feb 14, 2024Updated 2 years ago
- T5Voice is a lightweight PyTorch implementation of T5-based text-to-speech synthesis, supporting both streaming and non-streaming speech …☆28Nov 7, 2025Updated 9 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Target speaker automatic speech recognition (TS-ASR)☆15Oct 14, 2023Updated 2 years ago
- This repository presents an evaluation framework for speech-to-speech (S2S) models, following the methodology described in the EmphAsses …☆25Jan 9, 2024Updated 2 years ago
- Speaker-aware CTC (SACTC) for multi-talker overlapped speech recognition.☆22May 26, 2025Updated last year
- Legible, Scalable, Reproducible Foundation Models with Named Tensors and Jax☆16Jun 16, 2024Updated 2 years ago
- ☆12Mar 11, 2025Updated last year
- ☆13Sep 25, 2024Updated last year
- ☆34Jun 15, 2021Updated 5 years ago
- ☆28Aug 1, 2026Updated last week
- Speaker embedding for anime speech domain based on ECAPA_TDNN☆21Jun 22, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ESLTTS dataset☆16Feb 6, 2025Updated last year
- A minimalist Docker project to help people getting started with Node, WizardCoder, CTransformers, Python, Express and TypeScript. Ready t…☆14Jun 23, 2023Updated 3 years ago
- StyleTTS 2 Optimized Training Fork☆32Feb 2, 2025Updated last year
- Repository for speech paper reading☆33Aug 19, 2021Updated 4 years ago
- ☆33Feb 4, 2025Updated last year
- ☆14Oct 3, 2025Updated 10 months ago
- pytorch model for contexless-phoneme prediction from speech audio☆32Oct 30, 2025Updated 9 months ago
- [TASLP 2024] Textless Unit-to-Unit training for Many-to-Many Multilingual Speech-to-Speech Translation☆31Sep 6, 2024Updated last year
- Official implementation of the paper "Distilling a Pretrained Language Model to a Multilingual ASR Model" (Interspeech 2022)☆12Mar 12, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Inference code for Interspeech 2025 paper, "LSCodec: Low-Bitrate and Speaker-Decoupled Discrete Speech Codec"☆36Oct 23, 2025Updated 9 months ago
- Fully local voice interface for Claude Code on Apple Silicon. Parakeet STT + Kokoro TTS + SmartTurn EOU + dual VAD.☆33Mar 24, 2026Updated 4 months ago
- Official implementation of INTERSPECCH 2022 Radio2Speech: High Quality Speech Recovery from Radio Frequency Signals☆19Sep 19, 2025Updated 10 months ago
- Pybind11 bindings for Kaldi☆15Jul 11, 2026Updated 3 weeks ago
- A Benchmark Corpus for Low-Resource Cantonese Punctuation Restoration from Speech Transcripts☆15Dec 3, 2024Updated last year
- SpeechGLUE is a speech version of the GLUE benchmark, driven by text-to-speech.☆13Jun 2, 2023Updated 3 years ago
- Text-to-Speech Benchmark☆28Apr 2, 2026Updated 4 months ago