Offline streaming speech-to-text in the browser
☆27Aug 28, 2025Updated last year
Alternatives and similar repositories for wasm-speech-streaming
Users that are interested in wasm-speech-streaming are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Oct 11, 2024Updated last year
- Text-to-text alignment algorithm for speech recognition error analysis.☆34Jun 23, 2026Updated 2 months ago
- ☆29Sep 9, 2026Updated last week
- Open-weights voice acting pipeline combining zero-shot voice cloning with natural-language direction. Provide a reference voice (or gener…☆18May 25, 2026Updated 3 months ago
- ☆17Dec 18, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Parallelized automatic corpus collection for ASR. Forked from https://github.com/EgorLakomkin/KTSpeechCrawler☆23Mar 21, 2021Updated 5 years ago
- Forced alignment decoder for Whisper.☆16Mar 13, 2024Updated 2 years ago
- ☆19Mar 22, 2024Updated 2 years ago
- ☆20Sep 2, 2024Updated 2 years ago
- Train no-reference speech quality estimators with multiple datasets via learned, per-dataset alignments.☆19Aug 1, 2025Updated last year
- High-performance, semantic turn detection for conversational AI☆45Oct 1, 2025Updated 11 months ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- 🎵 muse: Music Separation☆11Feb 14, 2024Updated 2 years ago
- T5Voice is a lightweight PyTorch implementation of T5-based text-to-speech synthesis, supporting both streaming and non-streaming speech …☆28Nov 7, 2025Updated 10 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Target speaker automatic speech recognition (TS-ASR)☆15Oct 14, 2023Updated 2 years ago
- This repository presents an evaluation framework for speech-to-speech (S2S) models, following the methodology described in the EmphAsses …☆25Jan 9, 2024Updated 2 years ago
- Speaker-aware CTC (SACTC) for multi-talker overlapped speech recognition.☆22May 26, 2025Updated last year
- Legible, Scalable, Reproducible Foundation Models with Named Tensors and Jax☆16Jun 16, 2024Updated 2 years ago
- ☆12Mar 11, 2025Updated last year
- ☆13Sep 25, 2024Updated last year
- ☆34Jun 15, 2021Updated 5 years ago
- ☆28Aug 1, 2026Updated last month
- ESLTTS dataset☆16Feb 6, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A minimalist Docker project to help people getting started with Node, WizardCoder, CTransformers, Python, Express and TypeScript. Ready t…☆14Jun 23, 2023Updated 3 years ago
- StyleTTS 2 Optimized Training Fork☆32Feb 2, 2025Updated last year
- Repository for speech paper reading☆33Aug 19, 2021Updated 5 years ago
- Speaker embedding for anime speech domain based on ECAPA_TDNN☆23Jun 22, 2025Updated last year
- ☆33Feb 4, 2025Updated last year
- ☆14Oct 3, 2025Updated 11 months ago
- pytorch model for contexless-phoneme prediction from speech audio☆32Oct 30, 2025Updated 10 months ago
- [TASLP 2024] Textless Unit-to-Unit training for Many-to-Many Multilingual Speech-to-Speech Translation☆31Sep 6, 2024Updated 2 years ago
- Official implementation of the paper "Distilling a Pretrained Language Model to a Multilingual ASR Model" (Interspeech 2022)☆12Mar 12, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Inference code for Interspeech 2025 paper, "LSCodec: Low-Bitrate and Speaker-Decoupled Discrete Speech Codec"☆37Oct 23, 2025Updated 10 months ago
- Fully local voice interface for Claude Code on Apple Silicon. Parakeet STT + Kokoro TTS + SmartTurn EOU + dual VAD.☆35Mar 24, 2026Updated 5 months ago
- Official implementation of INTERSPECCH 2022 Radio2Speech: High Quality Speech Recovery from Radio Frequency Signals☆19Sep 19, 2025Updated 11 months ago
- Pybind11 bindings for Kaldi☆15Aug 20, 2026Updated 3 weeks ago
- A Benchmark Corpus for Low-Resource Cantonese Punctuation Restoration from Speech Transcripts☆15Dec 3, 2024Updated last year
- SpeechGLUE is a speech version of the GLUE benchmark, driven by text-to-speech.☆13Jun 2, 2023Updated 3 years ago
- Text-to-Speech Benchmark☆30Aug 18, 2026Updated last month