A library for real-time voice processing in web browsers
☆247Aug 11, 2026Updated this week
Alternatives and similar repositories for web-voice-processor
Users that are interested in web-voice-processor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- On-device streaming speech-to-text engine powered by deep learning☆670Updated this week
- On-device speech-to-text engine powered by deep learning☆483Updated this week
- On-device Speech-to-Intent engine powered by deep learning☆706Updated this week
- A module for normalising text.☆10Nov 6, 2019Updated 6 years ago
- On-device voice activity detection (VAD) powered by deep learning☆269Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A python tool that converts Arabic diacritised text to a sequence of phonemes and creates a pronunciation dictionary. This code is based …☆15Sep 5, 2017Updated 8 years ago
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- On-device voice assistant platform powered by deep learning☆705Apr 11, 2025Updated last year
- For IEEE ASRU(2025)☆15Jun 21, 2025Updated last year
- Web Component Knobs☆19Feb 28, 2023Updated 3 years ago
- A simple node.js MRCP (v.2) library☆11Oct 26, 2024Updated last year
- ☆33Feb 4, 2025Updated last year
- An upgrade framework for train and validate compare with icefall using Lightning.☆16Mar 26, 2025Updated last year
- Create modular, cross-browser, web audio pipelines to record and process audio in background threads. Comes with modules for VAD, ASR, re…☆48Apr 16, 2026Updated 3 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- SALT: STANDARDIZED AUDIO EVENT LABEL TAXONOMY☆16Nov 28, 2024Updated last year
- wake word spotting with kaldi☆19Dec 3, 2020Updated 5 years ago
- Lightweight Korean TTS Model based on FastSpeech2☆15Mar 4, 2026Updated 5 months ago
- On-device noise suppression powered by deep learning☆92Updated this week
- Tool for creating Kaldi nnet3 recipes using the International Phonetic Alphabet (IPA)☆10Jun 2, 2021Updated 5 years ago
- Fast & tiny DOM differ☆16Mar 14, 2024Updated 2 years ago
- SIGMORPHON 2020 Shared Task: Grapheme-to-Phoneme, Unsupervised Induction of Morphology, and Typologically Diverse Morphological Inflectio…☆36Apr 25, 2025Updated last year
- Speaker diarization benchmark framework☆45Jul 17, 2026Updated 3 weeks ago
- a typescript compiler for nestjs that does validation, openapi, sdk and more☆18Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Download and play English vocabulary's audio via command line.☆12Oct 1, 2020Updated 5 years ago
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IW…☆18Nov 30, 2022Updated 3 years ago
- On-device streaming text-to-speech engine powered by deep learning☆143Updated this week
- High-resolution real-time graphic audio spectrum analyzer JavaScript module with no dependencies.☆936Jul 19, 2026Updated 3 weeks ago
- Korean ASR Corpus generated from TEDx talks☆27Jan 11, 2019Updated 7 years ago
- a basic steganography library for png files☆13Apr 15, 2017Updated 9 years ago
- Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment☆14Feb 5, 2025Updated last year
- Converts Mandarin Chinese pinyin notation to IPA (international phonetic alphabet) notation☆19Nov 28, 2023Updated 2 years ago
- speech to text benchmark framework☆697Jul 8, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Hangul pronunciation and romanisation based on Wiktionary ko-pron lua module☆21Oct 17, 2018Updated 7 years ago
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challenge☆16Mar 26, 2022Updated 4 years ago
- ☆11Sep 5, 2025Updated 11 months ago
- Second SIGMORPHON Shared Task on Grapheme-to-Phoneme Conversions☆25Jun 7, 2021Updated 5 years ago
- A speech recognition library running in the browser thanks to a WebAssembly build of Vosk☆527Dec 7, 2025Updated 8 months ago
- 24-hour Automatic Speech Recognition☆27Jun 4, 2021Updated 5 years ago
- Baseline SWR implementation intended for use with Preact☆11Mar 6, 2022Updated 4 years ago