Batch Support for OpenAI Whisper
☆97Jan 19, 2024Updated 2 years ago
Alternatives and similar repositories for batch-whisper
Users that are interested in batch-whisper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Accelerate Whisper tasks such as transcription, by multiprocesing through parallelization☆25Oct 29, 2022Updated 3 years ago
- List of Podcast Feeds using iTunes API and script to download 6,000,000~ hours of English speech.☆31Apr 13, 2023Updated 3 years ago
- The wizard of oz code used for collecting goal-oriented dialogue systems☆13Oct 30, 2017Updated 8 years ago
- ☆39Jul 4, 2024Updated 2 years ago
- Speaker prediction for captions on the Lex Fridman podcast☆26Feb 14, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- openvino version of openai/whisper☆185Nov 6, 2023Updated 2 years ago
- Whisper finetuned on VinBigdata-VLSP2020-100h + KenLM☆38Oct 6, 2023Updated 2 years ago
- ☆10Dec 11, 2022Updated 3 years ago
- Transcription and Diarization based on OpenAI's Whisper☆25Aug 17, 2026Updated last month
- Various tools written or modified by me☆10Apr 14, 2026Updated 5 months ago
- Applies a query like mongo standards to knex query builder☆17May 18, 2020Updated 6 years ago
- ☆10Oct 17, 2021Updated 4 years ago
- A simple implementation for improving CosyVoice2 by GRPO method☆39May 5, 2026Updated 4 months ago
- Procedural island for A-Frame☆16May 5, 2017Updated 9 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- content.rdf.u8.gz☆11Dec 15, 2020Updated 5 years ago
- ☆13Sep 25, 2024Updated last year
- Enhanced Supertonic TTS with Docker, FastAPI, Web UI, and comprehensive API documentation☆22Dec 7, 2025Updated 9 months ago
- ☆10Nov 27, 2019Updated 6 years ago
- ☆23Oct 30, 2024Updated last year
- open-source Mandarian biased word dataset☆14Sep 21, 2023Updated 3 years ago
- Code for the Interspeech 2023 paper "A Joint Model for Pronunciation Assessment and Mispronunciation Detection and Diagnosis with Multi-t…☆25Nov 9, 2023Updated 2 years ago
- All awesome resources involving the awesome Nullstack☆13Jun 16, 2021Updated 5 years ago
- Slurm SPANK plugin to ease setup of SSH tunnels and port forwarding☆12Mar 21, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- just for fun☆14Mar 11, 2018Updated 8 years ago
- faster inference☆27Jan 20, 2025Updated last year
- ☆38Dec 26, 2022Updated 3 years ago
- Codenames AI☆12Jun 21, 2022Updated 4 years ago
- ☆17May 5, 2024Updated 2 years ago
- Musical mel transform for semi/quarter-tone features, written in ONNX-compatible PyTorch for audio AI neural networks☆21Feb 20, 2026Updated 7 months ago
- A Pytorch implemtentation of ICCV 2019 paper Face Swapping Gan (https://arxiv.org/abs/1908.05932)☆20Nov 11, 2019Updated 6 years ago
- A novel Inductive Logic Programming(ILP) system based on Meta Inverse Entailment in Python.☆21Jul 9, 2026Updated 2 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- State-of-the-art architecture for Plant Disease Detection using Deep Learning.☆10Jul 4, 2022Updated 4 years ago
- Official Implementation of EnCLAP (ICASSP 2024)☆96Jun 2, 2024Updated 2 years ago
- ☆11Sep 26, 2022Updated 3 years ago
- ☆12Nov 7, 2024Updated last year
- Compute WER and SER for speech recognition evaluation☆28Jun 6, 2026Updated 3 months ago
- A Colab Notebook for OpenAI Whisper and DeepL API, aiming to create human-comparable results of translation and transcription.☆33Feb 4, 2024Updated 2 years ago
- Text-To-Speech for NotebookLM☆39Jul 20, 2025Updated last year