Batch Support for OpenAI Whisper
☆97Jan 19, 2024Updated 2 years ago
Alternatives and similar repositories for batch-whisper
Users that are interested in batch-whisper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Accelerate Whisper tasks such as transcription, by multiprocesing through parallelization☆25Oct 29, 2022Updated 3 years ago
- List of Podcast Feeds using iTunes API and script to download 6,000,000~ hours of English speech.☆31Apr 13, 2023Updated 3 years ago
- ☆10Nov 12, 2024Updated last year
- The wizard of oz code used for collecting goal-oriented dialogue systems☆13Oct 30, 2017Updated 8 years ago
- ☆39Jul 4, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Prompting Large Language Models with Audio for General-Purpose Speech Summarization☆20May 14, 2025Updated last year
- Speaker prediction for captions on the Lex Fridman podcast☆26Feb 14, 2024Updated 2 years ago
- openvino version of openai/whisper☆184Nov 6, 2023Updated 2 years ago
- Whisper finetuned on VinBigdata-VLSP2020-100h + KenLM☆38Oct 6, 2023Updated 2 years ago
- Transcription and Diarization based on OpenAI's Whisper☆25Sep 9, 2025Updated 11 months ago
- ☆12Dec 10, 2022Updated 3 years ago
- Various tools written or modified by me☆10Apr 14, 2026Updated 3 months ago
- Expressive TTS Dataset for Assamese, Bengali, and Tamil.☆15Mar 6, 2025Updated last year
- ☆10Oct 17, 2021Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A simple implementation for improving CosyVoice2 by GRPO method☆39May 5, 2026Updated 3 months ago
- Procedural island for A-Frame☆16May 5, 2017Updated 9 years ago
- ☆13Sep 25, 2024Updated last year
- Enhanced Supertonic TTS with Docker, FastAPI, Web UI, and comprehensive API documentation☆21Dec 7, 2025Updated 8 months ago
- ☆11Apr 4, 2023Updated 3 years ago
- ☆23Oct 30, 2024Updated last year
- Slurm SPANK plugin to ease setup of SSH tunnels and port forwarding☆12Mar 21, 2024Updated 2 years ago
- faster inference☆27Jan 20, 2025Updated last year
- ☆38Dec 26, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Codenames AI☆12Jun 21, 2022Updated 4 years ago
- ☆17May 5, 2024Updated 2 years ago
- Musical mel transform for semi/quarter-tone features, written in ONNX-compatible PyTorch for audio AI neural networks☆20Feb 20, 2026Updated 5 months ago
- ☆24Dec 11, 2024Updated last year
- State-of-the-art architecture for Plant Disease Detection using Deep Learning.☆10Jul 4, 2022Updated 4 years ago
- Official Implementation of EnCLAP (ICASSP 2024)☆96Jun 2, 2024Updated 2 years ago
- ☆11Sep 26, 2022Updated 3 years ago
- ☆12Nov 7, 2024Updated last year
- Compute WER and SER for speech recognition evaluation☆28Jun 6, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A Colab Notebook for OpenAI Whisper and DeepL API, aiming to create human-comparable results of translation and transcription.☆33Feb 4, 2024Updated 2 years ago
- Text-To-Speech for NotebookLM☆39Jul 20, 2025Updated last year
- A dataset generator for family tree data.☆15Mar 17, 2020Updated 6 years ago
- Faster distil-whisper transcription with CTranslate2☆14Jan 23, 2024Updated 2 years ago
- generate granular word-level captions in srt format☆58Sep 26, 2022Updated 3 years ago
- A simple script to prepare dataset for training with TTS Tortoise model via https://git.ecker.tech/mrq/ai-voice-cloning☆12Jan 12, 2024Updated 2 years ago
- SpeechJudge: Towards Human-Level Judgment for Speech Naturalness (https://arxiv.org/abs/2511.07931)☆81Dec 23, 2025Updated 7 months ago