OpenAI's Whisper Audio to text transcription right into your web browser! An open source AI subtitling suite.
☆165Aug 27, 2023Updated 3 years ago
Alternatives and similar repositories for web-whisper
Users that are interested in web-whisper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Project for HIDING SPEAKER’S SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINE☆15Nov 30, 2022Updated 3 years ago
- ☆13May 1, 2026Updated 4 months ago
- The main repo for Stage Whisper — a free, secure, and easy-to-use transcription app for journalists, powered by OpenAI's Whisper automati…☆261Apr 27, 2023Updated 3 years ago
- This app is intended to automatically create a corpus for ASR systems using pseudo-labeling.☆27Feb 15, 2024Updated 2 years ago
- Using YouTube to prepare a speech recognition dataset for any language☆10Mar 30, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Finally, some decent sample sentences☆24Dec 3, 2023Updated 2 years ago
- Zero-shot Audio Classification using Whisper☆79Dec 12, 2022Updated 3 years ago
- A curated list of awesome OpenAI's Whisper☆107May 22, 2026Updated 3 months ago
- Accelerate Whisper tasks such as transcription, by multiprocesing through parallelization☆25Oct 29, 2022Updated 3 years ago
- A Go change detection server that can notify through various platforms☆96May 28, 2023Updated 3 years ago
- Pennyseed implemented with Next.js, Tailwindcss, Supabase, and Sendgrid☆18Jun 25, 2023Updated 3 years ago
- Enable RNNLM lattice rescoring with Pytorch [kaldi]☆12Jun 5, 2020Updated 6 years ago
- Korean read speech corpus (about 120 hours, 17GB) from National Institute of Korean Language☆43Feb 28, 2018Updated 8 years ago
- Wenet speech to text for react native☆10Nov 1, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Arabic Grapheme-to-Phoneme (G2P) Conversion☆15Mar 15, 2025Updated last year
- ☆27Jan 14, 2021Updated 5 years ago
- Implementation of OIDC and OAuth2☆17Jun 16, 2022Updated 4 years ago
- ☆16Jul 9, 2023Updated 3 years ago
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- Speechflow for emotion recognition related information decomposition☆10Jul 27, 2021Updated 5 years ago
- Dictionary of pairs of Korean word and IPA crawled from Wiktionary (Korean edition)☆24Nov 12, 2025Updated 9 months ago
- A GitHub authentication plugin for Node-RED☆19Aug 14, 2021Updated 5 years ago
- Streaming transcriber with whisper☆696May 1, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- A python script COMMAND LINE utility to AUTO GENERATE SUBTITLE FILE (using free Vosk Speech Recognition API) and TRANSLATED SUBTITLE FILE…☆10Updated this week
- Validate and Sign payload with JWT Protocol.☆17Dec 22, 2022Updated 3 years ago
- Aty-TTS: Improving fairness for spoken language understanding in atypical speech with Text-to-Speech☆12May 14, 2025Updated last year
- Speech recognition module for Python, supporting several engines and APIs, online and offline.☆13Mar 9, 2022Updated 4 years ago
- Robust Speech Recognition via Large-Scale Weak Supervision☆28Dec 16, 2023Updated 2 years ago
- ☆33Nov 27, 2021Updated 4 years ago
- Java Bindings for the C++ library DeepSpeech☆10Jun 4, 2020Updated 6 years ago
- Whisper as a Service (GUI and API with queuing for OpenAI Whisper)☆2,075Updated this week
- Transform audio-visual content into navigable knowledge.☆774Oct 25, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Node-red node to work with Mikrotik WiFi router☆16Jul 16, 2026Updated last month
- A python tool that converts Arabic diacritised text to a sequence of phonemes and creates a pronunciation dictionary. This code is based …☆15Sep 5, 2017Updated 9 years ago
- S3PRL for Speech Emotion Recognition (see s3prl > downstream)☆15Feb 28, 2026Updated 6 months ago
- An experiment of trying out whisper.cpp for real-time speech-to-text☆19Dec 25, 2022Updated 3 years ago
- A Neural Audio Codec (NAC) for Universal Audio☆47May 30, 2025Updated last year
- SERAB: a multi-lingual benchmark for speech emotion recognition☆28Dec 16, 2022Updated 3 years ago
- ☆13Oct 27, 2021Updated 4 years ago