Real-Time Whisper Voice Recognition with vosk model feedback.
☆120Jun 30, 2023Updated 3 years ago
Alternatives and similar repositories for vosper
Users that are interested in vosper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/☆834Sep 12, 2025Updated 10 months ago
- Streaming transcriber with whisper☆696May 1, 2023Updated 3 years ago
- A nearly-live implementation of OpenAI's Whisper, using sounddevice. Requires existing Whisper install.☆361Jul 20, 2025Updated last year
- A quick experiment to achieve almost realtime transcription using Whisper.☆186Sep 22, 2022Updated 3 years ago
- plugin manager for OpenVoiceOS , STT/TTS/Wakewords that can be used anywhere☆14Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆18Apr 28, 2021Updated 5 years ago
- Implementation of vocoders empowered with pytorch lightning☆18Jan 27, 2024Updated 2 years ago
- Real time speech to text transcription app.☆439Jan 14, 2023Updated 3 years ago
- Robust Speech Recognition via Large-Scale Weak Supervision☆91Aug 28, 2023Updated 2 years ago
- An extension of thu-spmi/CAT which contains a full-fledged implementation of CTC-CRF for Tensorflow.☆12Jul 5, 2021Updated 5 years ago
- Real time transcription with OpenAI Whisper.☆2,940Apr 15, 2025Updated last year
- Thin wrapper around OpenAI Whisper API with streaming support☆84Dec 5, 2025Updated 7 months ago
- WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)☆12Aug 1, 2025Updated 11 months ago
- ANDROID APP that can RECOGNIZE VLC LIVE AUDIO/VIDEO STREAMING (using free Android Developers Speech Recognition API) then TRANSLATE (usin…☆14May 5, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A free & open tool for transcribing audio interviews with offline ASR support☆25Dec 21, 2023Updated 2 years ago
- Voice memos recorded from the microphone, transcribed offline to text and converted to Joplin notes☆29Mar 1, 2024Updated 2 years ago
- A library for adding punctuation into a text from ASR.☆19May 8, 2023Updated 3 years ago
- A nearly-live implementation of OpenAI's Whisper.☆4,153Jul 17, 2026Updated last week
- ☆15Aug 25, 2022Updated 3 years ago
- Private voice keyboard, agent, AI chat, images, webcam, recordings, voice control with >= 4 GiB of VRAM.☆291Jun 16, 2026Updated last month
- Keyword Spotting for detecting a word in an audio file☆17Jul 21, 2019Updated 7 years ago
- Listen, transcribe, reply - Voice Assistant using OpenAI & ElevenLabs API's☆14Jun 24, 2023Updated 3 years ago
- Self hosted high quality voice recognition for de-googled Android using whisper. Like Siri or OK Google.☆74Dec 30, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Demo app showing video-powered customer service with Twilio for Python using the Flask web framework. Work in progress!☆16May 15, 2015Updated 11 years ago
- A python package to build AI-powered real-time audio applications☆2,005Jun 19, 2026Updated last month
- Audio-visual diarization pipeline used for creating VoxConverse dataset☆22Jun 6, 2025Updated last year
- This repository contains code for applying Data2Vec to pretrain Keyword Transformer model as described in "Improving Label-Deficient Keyw…☆32Mar 6, 2025Updated last year
- Simple, energy-based voice activity detection algorithm implementation.☆18Apr 21, 2024Updated 2 years ago
- Zero-shot multimodal punctuation insertion and truecasing using Whisper☆120Feb 4, 2023Updated 3 years ago
- Simple diarization model☆53Jun 13, 2025Updated last year
- Live transcription with OpenAi Whisper☆50Nov 11, 2022Updated 3 years ago
- Listen to any audio stream on your machine and print out the transcribed or translated audio.☆119Aug 16, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Go from raw audio files to a text-audio dataset automatically with OpenAI's Whisper.☆137Aug 14, 2023Updated 2 years ago
- ☆22Jun 30, 2021Updated 5 years ago
- 🐍 🤖 Pip installable package for StyleTTS 2 human-level text-to-speech and voice cloning☆159Jul 15, 2024Updated 2 years ago
- Evaluation script for VoxMovies dataset in PyTorch☆23Jan 12, 2024Updated 2 years ago
- Phonotate.App is a local, open-source Electron app built with React designed to simplify creating training data for StyleTTS 2 and voice …☆11Jan 17, 2025Updated last year
- StyleTTS 2 Optimized Training Fork☆32Feb 2, 2025Updated last year
- Official source for Catalan Language Models and resources made within Aina project.☆26Jul 28, 2023Updated 2 years ago