Project that allows one to use a microphone with OpenAI whisper.
☆787Jul 4, 2024Updated 2 years ago
Alternatives and similar repositories for whisper_mic
Users that are interested in whisper_mic are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Demo Programs for the "Talking Head(?) Anime from a Single Image 3: Now the Body Too" Project☆1,044Aug 29, 2023Updated 3 years ago
- A nearly-live implementation of OpenAI's Whisper, using sounddevice. Requires existing Whisper install.☆362Jul 20, 2025Updated last year
- Real time transcription with OpenAI Whisper.☆2,940Apr 15, 2025Updated last year
- Streaming transcriber with whisper☆696May 1, 2023Updated 3 years ago
- Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/☆835Sep 12, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Listen to any audio stream on your machine and print out the transcribed or translated audio.☆119Aug 16, 2023Updated 3 years ago
- AI Livestreamer for Youtube☆68Mar 19, 2023Updated 3 years ago
- Robust Speech Recognition via Large-Scale Weak Supervision☆108,593Aug 31, 2026Updated last week
- A nearly-live implementation of OpenAI's Whisper.☆4,252Updated this week
- Shared Voice Interface☆43Oct 21, 2023Updated 2 years ago
- Whisper realtime streaming for long speech-to-text transcription and translation☆3,671Nov 12, 2025Updated 9 months ago
- Whispering Tiger - OpenAI's whisper (and other models) with OSC and Websocket support. Allowing live transcription / translation in VRCha…☆540Updated this week
- Faster Whisper transcription with CTranslate2☆25,262Nov 19, 2025Updated 9 months ago
- A quick experiment to achieve almost realtime transcription using Whisper.☆186Sep 22, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Real time speech to text transcription app.☆440Jan 14, 2023Updated 3 years ago
- Robust Speech Recognition via Large-Scale Weak Supervision☆28Dec 16, 2023Updated 2 years ago
- Speaker prediction for captions on the Lex Fridman podcast☆26Feb 14, 2024Updated 2 years ago
- Real-time transcription using faster-whisper☆614Jul 23, 2024Updated 2 years ago
- AI Livestreamer for Youtube☆495Mar 18, 2023Updated 3 years ago
- Port of OpenAI's Whisper model in C/C++☆53,479Updated this week
- OpenAI Whisper ASR Webservice API☆3,331Aug 9, 2026Updated 3 weeks ago
- WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)☆23,916Aug 30, 2026Updated last week
- JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.☆4,681Apr 3, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆11Sep 5, 2025Updated last year
- Zero-shot multimodal punctuation insertion and truecasing using Whisper☆120Feb 4, 2023Updated 3 years ago
- ☆267Mar 19, 2023Updated 3 years ago
- openai/whisper + extra features☆89Oct 26, 2022Updated 3 years ago
- Real-Time Whisper Voice Recognition with vosk model feedback.☆120Jun 30, 2023Updated 3 years ago
- Project for HIDING SPEAKER’S SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINE☆15Nov 30, 2022Updated 3 years ago
- Minimal extension of OpenAI's Whisper adding speaker diarization with special tokens☆553Nov 6, 2023Updated 2 years ago
- Neural building blocks for speaker diarization: speech activity detection, speaker change detection, overlapped speech detection, speaker…☆10,514Updated this week
- The YouTube Text-To-Speech dataset is comprised of waveform audio extracted from YouTube videos alongside their English transcriptions☆53Apr 1, 2021Updated 5 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.☆4,116Jan 8, 2025Updated last year
- A multi-voice TTS system trained with an emphasis on quality☆14,873Nov 19, 2024Updated last year
- Multilingual Automatic Speech Recognition with word-level timestamps and confidence☆2,842Aug 17, 2026Updated 3 weeks ago
- Live transcription with OpenAi Whisper☆50Nov 11, 2022Updated 3 years ago
- 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production☆45,985Aug 16, 2024Updated 2 years ago
- Accelerate Whisper tasks such as transcription, by multiprocesing through parallelization☆25Oct 29, 2022Updated 3 years ago
- Thin wrapper around OpenAI Whisper API with streaming support☆84Dec 5, 2025Updated 9 months ago