☆267Mar 19, 2023Updated 3 years ago
Alternatives and similar repositories for stream-translator
Users that are interested in stream-translator are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Listen to any audio stream on your machine and print out the transcribed or translated audio.☆119Aug 16, 2023Updated 2 years ago
- A stream-translator fork with VAD based audio slicing & GPT / Gemini translation.☆211Updated this week
- faster-whisper livestream translation, OBS noise reduction, dual language subtitles☆82Apr 26, 2023Updated 3 years ago
- Wenet speech to text for react native☆10Nov 1, 2022Updated 3 years ago
- Project for HIDING SPEAKER’S SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINE☆15Nov 30, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 56 language, 1 model Multilingual ASR☆25Jul 25, 2021Updated 5 years ago
- Whisper realtime streaming for long speech-to-text transcription and translation☆3,657Nov 12, 2025Updated 8 months ago
- ☆13Dec 7, 2022Updated 3 years ago
- This app is intended to automatically create a corpus for ASR systems using pseudo-labeling.☆27Feb 15, 2024Updated 2 years ago
- ☆21Feb 20, 2023Updated 3 years ago
- Real time speech to text transcription app.☆439Jan 14, 2023Updated 3 years ago
- Faster Whisper transcription with CTranslate2☆24,595Nov 19, 2025Updated 8 months ago
- Whispering Tiger - OpenAI's whisper (and other models) with OSC and Websocket support. Allowing live transcription / translation in VRCha…☆536Jul 15, 2026Updated 2 weeks ago
- Real time transcription with OpenAI Whisper.☆2,939Apr 15, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- repo of files pertaining to realtime, offline translations using whisper realtime and argos translate. This repo is marked Creative Commo…☆19May 20, 2025Updated last year
- A nearly-live implementation of OpenAI's Whisper.☆4,190Updated this week
- Streaming transcriber with whisper☆696May 1, 2023Updated 3 years ago
- ☆16Jun 13, 2022Updated 4 years ago
- Data manipulation and transformation for audio signal processing, powered by PyTorch☆10Sep 30, 2024Updated last year
- Hed and supporting files for Chinese NNSVS Dataset Creation☆13Oct 14, 2025Updated 9 months ago
- Experimental code: sound file preprocessing to optimize Whisper transcriptions without hallucinated texts☆351Nov 12, 2024Updated last year
- One command to start a streaming ASR server.☆12Oct 2, 2024Updated last year
- An environment where you can try out faster-whisper immediately.☆37Nov 21, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Project that allows one to use a microphone with OpenAI whisper.☆788Jul 4, 2024Updated 2 years ago
- Enable RNNLM lattice rescoring with Pytorch [kaldi]☆12Jun 5, 2020Updated 6 years ago
- A python package to build AI-powered real-time audio applications☆2,006Jun 19, 2026Updated last month
- sherpa with mlx☆15Aug 2, 2025Updated 11 months ago
- This repo contains the code for the tutorial for using the CrewAI agent framework to generate Sales Reports based on Salesforce data☆13Mar 16, 2024Updated 2 years ago
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/☆834Sep 12, 2025Updated 10 months ago
- A enterprise-grade Chinese-English code switch punctuator from funasr.☆34Apr 26, 2024Updated 2 years ago
- Chrome/Edge BROWSER EXTENSION that can RECOGNIZE any live audio/video streaming then TRANSLATE it for FREE (using unofficial online Googl…☆121Jul 28, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This is a project focused on Faster Whisper, a streaming speech recognition project.☆18Sep 27, 2024Updated last year
- Whisper combined with Silero VAD, for improved long-form transcriptions☆55Dec 11, 2022Updated 3 years ago
- Zero-shot multimodal punctuation insertion and truecasing using Whisper☆120Feb 4, 2023Updated 3 years ago
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- ☆15Jun 4, 2021Updated 5 years ago
- Speechflow for emotion recognition related information decomposition☆10Jul 27, 2021Updated 5 years ago
- Real-time transcription using faster-whisper☆615Jul 23, 2024Updated 2 years ago