Data preparation utility for the finetuning of OpenAI's Whisper model.
☆18Sep 17, 2026Updated this week
Alternatives and similar repositories for whisper-prep
Users that are interested in whisper-prep are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository contains code for fine-tuning the Whisper speech-to-text model.☆26Updated this week
- Redesign of the new version of the QB-Inventory☆13Feb 19, 2024Updated 2 years ago
- Video Summarization Transformer: Implementation in PyTorch of the Transformer model for video summarisation☆10Oct 27, 2020Updated 5 years ago
- Docker XPRA HTML5 Image with OpenGL support for NVIDIA cards☆13Oct 28, 2020Updated 5 years ago
- A REST API for neural machine translation☆14Jun 11, 2026Updated 3 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Generating artificial disfluencies from fluent text easily and promptly☆15Sep 28, 2022Updated 3 years ago
- Enhanced Reverberation As Supervision (ERAS) for unsupervised reverberant speech separation☆15Aug 1, 2024Updated 2 years ago
- The implementation of MDNet, which is in submission to Interspeech2022☆14May 1, 2022Updated 4 years ago
- Multi Stage Attentional UNet☆11Dec 23, 2021Updated 4 years ago
- ☆18Oct 16, 2018Updated 7 years ago
- A Range-Null Space Decomposition Approach for Fast and Flexible Spectral Compressive Imaging☆11May 18, 2023Updated 3 years ago
- End-to-End Speech Processing Toolkit☆12Nov 16, 2024Updated last year
- Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment☆14Feb 5, 2025Updated last year
- ☆10Jun 11, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- "Unsupervised Paraphrase Generation using Pre-trained Language Model."☆22Aug 28, 2020Updated 6 years ago
- Example of a self-updating application using tufup.☆20Oct 3, 2025Updated 11 months ago
- Python script to transform the Mobile Detect JSON database into an UA-based mobile detection VCL subroutine easily integrable in any Varn…☆14Nov 13, 2023Updated 2 years ago
- Waste images dataset☆17Apr 29, 2017Updated 9 years ago
- Avalinguo Audio Dataset: Dataset for Speaker Fluency Level Classification☆13Aug 13, 2018Updated 8 years ago
- Phonetically-Oriented Word Error Rate☆36May 4, 2019Updated 7 years ago
- ☆12May 10, 2023Updated 3 years ago
- ☆17Mar 30, 2023Updated 3 years ago
- docker container to steam via OBS, NDI, NVENC support☆25Sep 14, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Triton backend for https://github.com/OpenNMT/CTranslate2☆11Aug 20, 2024Updated 2 years ago
- ☆16Apr 24, 2021Updated 5 years ago
- This small project demonstrates how to integrate WordPress blog entries into queries for a RAG-based (Retriever-Augmented Generation) lan…☆11Apr 2, 2024Updated 2 years ago
- ☆11Oct 14, 2023Updated 2 years ago
- Whisper Speaker Identification (WSI), a cutting-edge model for multilingual speaker identification.☆27Jun 29, 2026Updated 2 months ago
- ☆16Sep 19, 2024Updated 2 years ago
- Noise-Aware Speech Separation with Contrastive Learning☆21Apr 25, 2024Updated 2 years ago
- An Interactive Tool for Annotating Discourse Structure and Text Improvement☆16Sep 15, 2021Updated 5 years ago
- Generated Audio Samples by ALGAN-VC model are available in the folder☆18Feb 25, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- MT Evaluation in Many Languages via Zero-Shot Paraphrasing☆102Jul 25, 2024Updated 2 years ago
- ☆14Nov 13, 2025Updated 10 months ago
- Noise15 , Noisex-92 and Nonspeech☆58Nov 17, 2020Updated 5 years ago
- Voice conversion using deep adversarial learning☆17Oct 29, 2021Updated 4 years ago
- ☆16Jul 23, 2024Updated 2 years ago
- Transfer learning approach to pronunciation scoring☆12Jan 17, 2024Updated 2 years ago
- [WIP] Scripts for fine-tuning Whisper☆221Jul 2, 2026Updated 2 months ago