SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition
☆96Sep 5, 2020Updated 5 years ago
Alternatives and similar repositories for SpecAugment
Users that are interested in SpecAugment are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Implementation of SpecAugment with Tensorflow & Pytorch, introduced by Google Brain☆655Apr 5, 2022Updated 4 years ago
- ☆13Jun 2, 2022Updated 4 years ago
- ☆18Apr 12, 2021Updated 5 years ago
- Learnable Gammatone Filterbank (LGTFB) and Equal-loudness Normalization (EN)☆13Apr 24, 2020Updated 6 years ago
- ☆14Jun 26, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Jan 10, 2024Updated 2 years ago
- A python implementation of a traditional Dynamic Range Compressor☆14Oct 30, 2020Updated 5 years ago
- ☆17Aug 9, 2024Updated 2 years ago
- Reproduction of a paper"Small-footprint keyword spotting using deep neural networks"☆12Mar 11, 2019Updated 7 years ago
- ☆18Nov 15, 2021Updated 4 years ago
- fast SpecAugmentation code with numpy and scipy☆31Jul 5, 2019Updated 7 years ago
- Code for ICLR 2024 Paper: CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models☆23Jul 10, 2024Updated 2 years ago
- ☆12Jul 14, 2024Updated 2 years ago
- An official documentation of the paper <Wave-U-Mamba: An End-To-End Framework For High-Quality And Efficient Speech Super Resolution>.☆26Oct 29, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Data generator for creating synthetic audio mixtures suitable for DCASE Challenge 2022 Task 3☆47Apr 5, 2023Updated 3 years ago
- creating audio preprocessing features in TensorFlow keras layers,☆14Jul 13, 2021Updated 5 years ago
- 🔦 A Pytorch implementation of GoogleBrain's SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition☆501Jun 11, 2021Updated 5 years ago
- tf 2.0 implementation of Listen, attend and spell☆21Jan 19, 2021Updated 5 years ago
- FINALLY: Fast and universal speech enhancement model delivering studio-quality audio for a wide range of recordings.☆28Apr 1, 2026Updated 4 months ago
- Automatic Speech Recognition with TensorFlow(CNN+BLSTM+CTC)☆12Aug 9, 2018Updated 8 years ago
- 🔊 A comprehensive list of open-source datasets for voice and sound computing (50+ datasets).☆20Apr 1, 2021Updated 5 years ago
- Download and preperation tool for free speech corpora.☆16Apr 28, 2019Updated 7 years ago
- Chainer implementation of between-class learning for sound recognition https://arxiv.org/abs/1711.10282☆95Mar 27, 2018Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code for the AAAI 2022 paper "SSAST: Self-Supervised Audio Spectrogram Transformer".☆430Aug 14, 2022Updated 3 years ago
- ☆14Oct 14, 2020Updated 5 years ago
- Implementation of the paper "Self-supervised Learning with Random-projection Quantizer for Speech Recognition" in Pytorch.☆97May 25, 2023Updated 3 years ago
- Code for the paper "MULTI-BAND MASKING FOR WAVEFORM-BASED SINGING VOICE SEPARATION" that was accepted on EUSIPCO2022☆15Jun 18, 2022Updated 4 years ago
- Group review spammer detection☆10Sep 9, 2019Updated 6 years ago
- Temporally-aligned Audio CaptiOnS for Language-Audio Pretraining☆16Oct 12, 2025Updated 10 months ago
- Project page of "2026-ICLR Echo: Towards Advanced Audio Comprehension via Audio-Interleaved Reasoning"☆16Mar 26, 2026Updated 4 months ago
- Keyword spotting for audio with attention (KWS model for audio)☆18Jul 15, 2021Updated 5 years ago
- The official repository TimeAudio, a comprehensive framework that incorporates fine-grained acoustic cues into LALMs with enhanced module…☆31Nov 18, 2025Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Tools for speech processing, keyword spotting☆16Mar 11, 2020Updated 6 years ago
- (ICASSP 2024) Official Implementation of "Stethoscope-guided Supervised Contrastive Learning for Cross-domin Adaptation on Respiratory So…☆19Dec 5, 2024Updated last year
- Use speech_to_text for keyword search in audio files.☆12May 5, 2021Updated 5 years ago
- (NeurIPS 2023 Workshop on DGM4H) Official Implementation of "Adversarial Fine-tuning using Generated Respiratory Sound to Address Class I…☆19Dec 5, 2024Updated last year
- Listen, Attend and Spell (LAS) framework for speech recognition (see https://arxiv.org/pdf/1508.01211.pdf).☆32Jun 27, 2019Updated 7 years ago
- Improved Speech Enhancement GANs☆13Jun 24, 2020Updated 6 years ago
- A deep learning approach for respiratory audio discovery and classification.☆14Sep 30, 2024Updated last year