π΅ A repository for manually annotating files to create labeled acoustic datasets for machine learning.
β47Feb 20, 2022Updated 4 years ago
Alternatives and similar repositories for sound_event_detection
Users that are interested in sound_event_detection are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Language modelling for sound event detectionβ20Jan 2, 2020Updated 6 years ago
- Python library for rapid prototyping of environmental sound analysis systemsβ44May 20, 2022Updated 4 years ago
- Easy to use Audio Tagging in PyTorchβ23Aug 22, 2021Updated 4 years ago
- Reading list for research topics in Sound AIβ200Aug 8, 2024Updated 2 years ago
- Using YouTube to prepare a speech recognition dataset for any languageβ10Mar 30, 2021Updated 5 years ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- β26Updated this week
- β16Apr 11, 2019Updated 7 years ago
- Documentation of the Two!Ears Auditory Modelβ13Feb 14, 2019Updated 7 years ago
- Official Implementation of DMT: Dual Mean-Teacher in PyTorch.β10Oct 27, 2023Updated 2 years ago
- Permutation invariant training in PyTorchβ13Oct 2, 2020Updated 5 years ago
- Read audio with FFmpeg into NumPy/PyTorch via ctypes (standard library module)β11Aug 12, 2020Updated 6 years ago
- Detect emotion from audioβ14Nov 20, 2018Updated 7 years ago
- A recipe for creating a Speaker Identification system built on Kaldi.β15Jan 2, 2020Updated 6 years ago
- Filter Bank Implementaion as Convolutional Neural Network using Python Kerasβ17Dec 18, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Adapt Kaldi-ASR nnet3 chain models from Zamia-Speech.org to a different language modelβ33Jan 26, 2020Updated 6 years ago
- β22Jun 30, 2021Updated 5 years ago
- Evaluation script for VoxMovies dataset in PyTorchβ23Jan 12, 2024Updated 2 years ago
- My system for the DCASE 2022 Task 3 Sound Event Localizaiton and Detection.β12Nov 12, 2022Updated 3 years ago
- Development Toolkit for the VoxCeleb Speaker Recognition Challenge 2020β43Jul 17, 2020Updated 6 years ago
- Filtering and Noise Adding Toolβ29May 27, 2022Updated 4 years ago
- [ICLR 2022] "Audio Lottery: Speech Recognition Made Ultra-Lightweight, Noise-Robust, and Transferable", by Shaojin Ding, Tianlong Chen, Zβ¦β32Apr 8, 2022Updated 4 years ago
- A tool/script for batch speech data enhancement with speed/volume/RIRS/MUSANβ25Jun 28, 2020Updated 6 years ago
- Sound event detection with depthwise separable and dilated convolutions.β53Mar 30, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A Diffrentiable WFST-based End-to-End Automatic Speech Recognition toollkit with flexible topology supportβ12Feb 15, 2026Updated 6 months ago
- The active learning algorithm, mismatch-first farthest-traversal. Implementation and visualization.β12Dec 25, 2021Updated 4 years ago
- In this repository, I try to combine k2 with speechbrain to decode well and fastly.β16Jun 17, 2022Updated 4 years ago
- π Easily apply audio-related machine learning models trained on the AudioSet dataset (527+ models/classes).β31Jun 17, 2024Updated 2 years ago
- Comprehensive Python library for speech and voice.β32Dec 8, 2022Updated 3 years ago
- β55Oct 17, 2023Updated 2 years ago
- Echo aware source separationβ13May 29, 2018Updated 8 years ago
- Vim Speech Recognition Experimentsβ20May 30, 2025Updated last year
- Support for Clarity Enhancement and Prediction Challenges (obsolete - see README)β48Apr 14, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Learning Domain-Invariant Transformation for Speaker Verification.β11Jun 13, 2023Updated 3 years ago
- Pypi installable TDNN and TDNN-F layers for PyTorch based acoustic model trainingβ41Dec 18, 2020Updated 5 years ago
- MirasVoice is a data set consisting speech samples from bilinguals to train neural network for optimization of speaker verification algorβ¦β19Mar 15, 2020Updated 6 years ago
- This repository describes our reproducible framework for assessing self-supervised representation learning from speechβ52Oct 8, 2021Updated 4 years ago
- Benchmark popular audio i/o packagesβ152Dec 19, 2023Updated 2 years ago
- Single and multichannel sound event detection using convolutional recurrent neural networks. DCASE 2017 real-life sound event detection wβ¦β202Jun 21, 2022Updated 4 years ago
- Artie Bias Corpus: an audio corpus + code for detecting demographic biasβ20Jul 21, 2020Updated 6 years ago