Sound augmentation using Large-scale audio dataset (Audioset)
☆45Jun 29, 2021Updated 5 years ago
Alternatives and similar repositories for audioset_augmentor
Users that are interested in audioset_augmentor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆21Jan 13, 2020Updated 6 years ago
- A simple tutorial on setting up Sparrowhawk - a text-to-speech normalization engine☆14Oct 16, 2017Updated 8 years ago
- BurrMill core☆22Nov 2, 2021Updated 4 years ago
- This is a mirror of https://gitlab.com/tiro-is/tiro-speech-core☆15Jun 19, 2023Updated 3 years ago
- readers that enable reading kaldi ark in tensorflow☆17Mar 7, 2018Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Sound Related Deep Learning Tasks boosting repository with pytorch☆88Jul 25, 2024Updated 2 years ago
- Train a Deep Learning model to classify audio embeddings on IBM's Deep Learning as a Service (DLaaS) platform - Watson Machine Learning☆102Sep 17, 2025Updated 10 months ago
- ☆17Apr 14, 2023Updated 3 years ago
- Tools for ASR Corpus Generation from Online Video☆140Feb 10, 2019Updated 7 years ago
- Deep learning based speech source separation using Pytorch☆319Nov 20, 2020Updated 5 years ago
- A collection of basic python modules for spoken natural language processing☆55Dec 1, 2019Updated 6 years ago
- 📁 This repo makes it easy to download the raw audio files from AudioSet (32.45 GB, 632 classes).☆106Aug 1, 2023Updated 2 years ago
- ☆45Oct 24, 2020Updated 5 years ago
- ☆13Mar 25, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- AudioVisual Diarization - Supervised and Unsupervised☆15Nov 22, 2022Updated 3 years ago
- Attacking Speaker Recognition with Deep Generative Models☆34Mar 24, 2023Updated 3 years ago
- speech engine training projects☆29Apr 19, 2021Updated 5 years ago
- Phonetically-Oriented Word Error Rate☆36May 4, 2019Updated 7 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- Segment a given audio into utterances using a trained end-to-end ASR model.☆75Oct 9, 2020Updated 5 years ago
- Convolutional Neural Network for multitrack mix leveling☆19Jun 25, 2018Updated 8 years ago
- ☆17Feb 14, 2020Updated 6 years ago
- An audio classification system for learning with out-of-distribution data☆33Dec 8, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SOTA punctation restoration (for e.g. automatic speech recognition) deep learning model based on BERT pre-trained model☆182May 17, 2019Updated 7 years ago
- ☆75Jan 6, 2020Updated 6 years ago
- ☆229Feb 9, 2020Updated 6 years ago
- Some useful features of speech process, such as MFCC, gammatone filterbank, GFCC, spectrum(power spectrum and log-power spectrum), Amplit…☆129Aug 12, 2020Updated 5 years ago
- steps to perform text-based speaker diarization with kaldi toolkit☆12Nov 2, 2018Updated 7 years ago
- A database of number names for 186 languages, locales, and scripts☆67Mar 3, 2023Updated 3 years ago
- ☆76Mar 18, 2022Updated 4 years ago
- 🏥 🎤 The largest clinical study in the world to collect voice data labeled with health information (N>6,000 participants, 48 utterances…☆32Apr 2, 2025Updated last year
- RawNet: Fast End-to-End Neural Vocoder☆43May 29, 2019Updated 7 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code for Speaker Change Detection in Broadcast TV using Bidirectional Long Short-Term Memory Networks☆67Jul 14, 2020Updated 6 years ago
- ☆22Sep 24, 2018Updated 7 years ago
- ASR with PyTorch☆139Mar 10, 2019Updated 7 years ago
- ☆17May 31, 2023Updated 3 years ago
- 语音切割,python ,webrtc☆11Sep 28, 2018Updated 7 years ago
- Repository for our Interspeech2020 general-purpose voice activity detection (GPVAD) paper☆141Aug 3, 2023Updated 2 years ago
- ☆131Aug 9, 2018Updated 7 years ago