This UI serves as a Synthetic ASR Dataset Generator powered by/for OpenAI Whisper, enabling users to capture audio, transcribing it, on the fly and manage the generated dataset 🤗. Fine tune Whisper or enhanced and custom datasets
☆34Nov 26, 2024Updated last year
Alternatives and similar repositories for Whisper-Synthetic-ASR-Dataset-Generator
Users that are interested in Whisper-Synthetic-ASR-Dataset-Generator are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Text-based media editing interface☆16Aug 9, 2017Updated 8 years ago
- A fast CPU-first video/audio transcriber for generating caption files with Whisper and CTranslate2, hosted on Hugging Face Spaces.☆11Updated this week
- A Python package for converting numbers expressed in natural language to numerical values.☆13Nov 25, 2023Updated 2 years ago
- ☆34May 15, 2023Updated 3 years ago
- Can Neural Networks reconstruct missing audio data? What about GANs?☆18Nov 6, 2019Updated 6 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Analyze and visualize how rhythm, timbre, loudness, pitch, spectral characteristics and other key audio features evolve over time across …☆11Updated this week
- Speech-to-text transcription VST3/ARA plugin☆62Jun 8, 2026Updated last month
- Demo how to use use Unix Domain Sockets in Swift on macOS.☆13Sep 24, 2021Updated 4 years ago
- ComfyUI port of SDWebUI Vectorscope CC and Diffusion CG extensions☆21Feb 24, 2025Updated last year
- Custom node for ComfyUI. Add a node for drawing text to the area of SEGS.☆14Mar 30, 2025Updated last year
- Voice activity detection and speaker gender segmentation audiovisual corpus☆16Jan 20, 2025Updated last year
- Human body part segmentation model, trained with 22 class labels.☆17Sep 28, 2023Updated 2 years ago
- Single Image Haze Removal Using AODNet in Pytorch☆15Mar 5, 2021Updated 5 years ago
- A Java port of whisper 3, based on the huggingface version, using DJL.☆17Apr 3, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆10Dec 10, 2021Updated 4 years ago
- Comparing Audio Features for Unsupervised Sound Classification☆10Jun 22, 2022Updated 4 years ago
- Whisfusion: Parallel ASR Decoding via a Diffusion Transformer☆31Aug 22, 2025Updated 11 months ago
- Trainer and Evaluation scripts for fine-tuning Whisper models for the Ukrainian language☆23Jan 13, 2023Updated 3 years ago
- ☆10Aug 3, 2019Updated 6 years ago
- Transform audio files into mel spectrograms for text-to-speech model training☆12Aug 25, 2021Updated 4 years ago
- ☆12May 1, 2019Updated 7 years ago
- Official source code for the paper "Tailored Design of Audio-Visual Speech Recognition Models using Branchformers"☆15Feb 24, 2025Updated last year
- Scripts to convert audio files to spectrograms and back☆12Nov 23, 2017Updated 8 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A utility to fetch and display dns names from the SSL/TLS cert data☆15Aug 11, 2023Updated 2 years ago
- ☆13Aug 25, 2021Updated 4 years ago
- Resnet50 Quantization for Inference Speedup in PyTorch☆23Jan 30, 2021Updated 5 years ago
- Collaborative audio annotation tool☆17Sep 16, 2022Updated 3 years ago
- Encode an image to sound (WAV file) and view it as a spectrogram. Optimized Python 3 version.☆11Jan 25, 2023Updated 3 years ago
- Prepare spectrograms from audio for training a Riffusion model☆16Mar 6, 2023Updated 3 years ago
- Rainbowgram with Python☆13Jan 28, 2019Updated 7 years ago
- ☆19May 9, 2019Updated 7 years ago
- Arduino/AVR C code for controlling the MOS6581 SID sound chip over MIDI☆11Oct 14, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆28Jun 28, 2024Updated 2 years ago
- Keras implementation of conditional waveGAN. Application to knocking sound effects with emotion.☆10Jun 22, 2020Updated 6 years ago
- Grad-CAM (Gradient-weighted Class Activation Mapping)☆13Dec 20, 2019Updated 6 years ago
- Convert images to audio for display in a spectrogram☆13Apr 17, 2018Updated 8 years ago
- CNN-to-FPGA-framework for small CNN, written in VHDL and Python☆24Jun 8, 2021Updated 5 years ago
- 95.76% on CIFAR-10 with TensorFlow2☆32Oct 21, 2021Updated 4 years ago
- RPi program to use Bluetooth and/or USB gamepads and mice on retro 8/16-bit computers (C64, Amiga, etc)☆15Dec 11, 2020Updated 5 years ago