Real-Time Whisper Voice Recognition with vosk model feedback.
☆120Jun 30, 2023Updated 3 years ago
Alternatives and similar repositories for vosper
Users that are interested in vosper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/☆834Sep 12, 2025Updated 11 months ago
- Streaming transcriber with whisper☆696May 1, 2023Updated 3 years ago
- A nearly-live implementation of OpenAI's Whisper, using sounddevice. Requires existing Whisper install.☆362Jul 20, 2025Updated last year
- A quick experiment to achieve almost realtime transcription using Whisper.☆186Sep 22, 2022Updated 3 years ago
- plugin manager for OpenVoiceOS , STT/TTS/Wakewords that can be used anywhere☆14Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆18Apr 28, 2021Updated 5 years ago
- React hook for OpenAI Whisper with speech recorder, real-time transcription, and silence removal built-in☆785Apr 30, 2024Updated 2 years ago
- Implementation of vocoders empowered with pytorch lightning☆18Jan 27, 2024Updated 2 years ago
- Real time speech to text transcription app.☆439Jan 14, 2023Updated 3 years ago
- Robust Speech Recognition via Large-Scale Weak Supervision☆91Aug 28, 2023Updated 2 years ago
- An extension of thu-spmi/CAT which contains a full-fledged implementation of CTC-CRF for Tensorflow.☆12Jul 5, 2021Updated 5 years ago
- Real time transcription with OpenAI Whisper.☆2,940Apr 15, 2025Updated last year
- Thin wrapper around OpenAI Whisper API with streaming support☆84Dec 5, 2025Updated 8 months ago
- Jupyter notebooks for PuLID face transfer with Flux.1 dev. Able to run on Google Colab Free Tier☆18Dec 18, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ANDROID APP that can RECOGNIZE VLC LIVE AUDIO/VIDEO STREAMING (using free Android Developers Speech Recognition API) then TRANSLATE (usin…☆14May 5, 2024Updated 2 years ago
- tkinter desktop chat interface with OpenAI's gpt-3.5-turbo API☆11Apr 29, 2023Updated 3 years ago
- A free & open tool for transcribing audio interviews with offline ASR support☆25Dec 21, 2023Updated 2 years ago
- Voice memos recorded from the microphone, transcribed offline to text and converted to Joplin notes☆29Mar 1, 2024Updated 2 years ago
- Speaker diarization service☆28Jul 2, 2026Updated last month
- A library for adding punctuation into a text from ASR.☆20May 8, 2023Updated 3 years ago
- A nearly-live implementation of OpenAI's Whisper.☆4,228Aug 4, 2026Updated last week
- PolEval 2021 Task 1☆15Jun 28, 2022Updated 4 years ago
- ☆15Aug 25, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Private voice keyboard, agent, AI chat, images, webcam, recordings, voice control with >= 4 GiB of VRAM.☆293Updated this week
- Keyword Spotting for detecting a word in an audio file☆17Jul 21, 2019Updated 7 years ago
- Listen, transcribe, reply - Voice Assistant using OpenAI & ElevenLabs API's☆14Jun 24, 2023Updated 3 years ago
- Self hosted high quality voice recognition for de-googled Android using whisper. Like Siri or OK Google.☆74Dec 30, 2023Updated 2 years ago
- IPsec Tunneling☆11Nov 30, 2022Updated 3 years ago
- A python package to build AI-powered real-time audio applications☆2,015Jun 19, 2026Updated last month
- Audio-visual diarization pipeline used for creating VoxConverse dataset☆22Jun 6, 2025Updated last year
- This repository contains code for applying Data2Vec to pretrain Keyword Transformer model as described in "Improving Label-Deficient Keyw…☆32Mar 6, 2025Updated last year
- Zero-shot multimodal punctuation insertion and truecasing using Whisper☆120Feb 4, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Simple diarization model☆53Jun 13, 2025Updated last year
- ☆33Aug 6, 2026Updated last week
- Listen to any audio stream on your machine and print out the transcribed or translated audio.☆119Aug 16, 2023Updated 3 years ago
- Go from raw audio files to a text-audio dataset automatically with OpenAI's Whisper.☆137Aug 14, 2023Updated 3 years ago
- ☆22Jun 30, 2021Updated 5 years ago
- 🐍 🤖 Pip installable package for StyleTTS 2 human-level text-to-speech and voice cloning☆159Jul 15, 2024Updated 2 years ago
- Evaluation script for VoxMovies dataset in PyTorch☆23Jan 12, 2024Updated 2 years ago