A library for real-time voice processing in web browsers
☆247Sep 3, 2026Updated this week
Alternatives and similar repositories for web-voice-processor
Users that are interested in web-voice-processor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- benchmark for Speech-to-Intent engines☆18Jul 30, 2026Updated last month
- Picovoice Browser Extension☆17Jun 24, 2026Updated 2 months ago
- On-device streaming speech-to-text engine powered by deep learning☆671Updated this week
- On-device speech-to-text engine powered by deep learning☆483Updated this week
- On-device Speech-to-Intent engine powered by deep learning☆710Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for the ⛅ weather and 🎥 movies skills☆13Jan 15, 2019Updated 7 years ago
- A module for normalising text.☆10Nov 6, 2019Updated 6 years ago
- On-device voice activity detection (VAD) powered by deep learning☆270Updated this week
- On-device speaker diarization powered by deep learning☆77Updated this week
- A python tool that converts Arabic diacritised text to a sequence of phonemes and creates a pronunciation dictionary. This code is based …☆15Sep 5, 2017Updated 9 years ago
- On-device wake word detection powered by deep learning☆4,931Updated this week
- For IEEE ASRU(2025)☆15Jun 21, 2025Updated last year
- Web Component Knobs☆19Feb 28, 2023Updated 3 years ago
- A simple node.js MRCP (v.2) library☆11Oct 26, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- DeepSpeechNotes is a note taking app using Mozilla's DeepSpeech technology to transcribe speech into text notes.☆18Jan 6, 2023Updated 3 years ago
- Text-to-Speech Benchmark☆29Aug 18, 2026Updated 2 weeks ago
- Create modular, cross-browser, web audio pipelines to record and process audio in background threads. Comes with modules for VAD, ASR, re…☆48Apr 16, 2026Updated 4 months ago
- SALT: STANDARDIZED AUDIO EVENT LABEL TAXONOMY☆16Nov 28, 2024Updated last year
- wake word spotting with kaldi☆19Dec 3, 2020Updated 5 years ago
- Lightweight Korean TTS Model based on FastSpeech2☆15Mar 4, 2026Updated 6 months ago
- On-device noise suppression powered by deep learning☆94Updated this week
- Tool for creating Kaldi nnet3 recipes using the International Phonetic Alphabet (IPA)☆10Jun 2, 2021Updated 5 years ago
- Fast & tiny DOM differ☆16Mar 14, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- SIGMORPHON 2020 Shared Task: Grapheme-to-Phoneme, Unsupervised Induction of Morphology, and Typologically Diverse Morphological Inflectio…☆36Apr 25, 2025Updated last year
- Adobe XD Cloud Content API samples☆18Mar 3, 2023Updated 3 years ago
- CLASP: Contrastive Language-Speech Pretraining for Multilingual Multimodal Information Retrieval☆13Jun 27, 2025Updated last year
- Download and play English vocabulary's audio via command line.☆12Oct 1, 2020Updated 5 years ago
- On-device streaming text-to-speech engine powered by deep learning☆145Updated this week
- Me building a simple synthesizer and sequencer to learn about Web Audio. Check out the wiki.☆16Oct 11, 2021Updated 4 years ago
- High-resolution real-time graphic audio spectrum analyzer JavaScript module with no dependencies.☆946Jul 19, 2026Updated last month
- Korean ASR Corpus generated from TEDx talks☆27Jan 11, 2019Updated 7 years ago
- a basic steganography library for png files☆13Apr 15, 2017Updated 9 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment☆14Feb 5, 2025Updated last year
- Pybind11 bindings for Kaldi☆15Aug 20, 2026Updated 2 weeks ago
- speech to text benchmark framework☆697Jul 8, 2026Updated last month
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challenge☆16Mar 26, 2022Updated 4 years ago
- A speech recognition library running in the browser thanks to a WebAssembly build of Vosk☆529Dec 7, 2025Updated 8 months ago
- 24-hour Automatic Speech Recognition☆27Jun 4, 2021Updated 5 years ago
- A simple web app to record audio and a Dialogflow agent to playback the audio as an Action.☆22May 16, 2019Updated 7 years ago