π΅ A repository for manually annotating files to create labeled acoustic datasets for machine learning.
β47Feb 20, 2022Updated 4 years ago
Alternatives and similar repositories for sound_event_detection
Users that are interested in sound_event_detection are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Language modelling for sound event detectionβ20Jan 2, 2020Updated 6 years ago
- Python library for rapid prototyping of environmental sound analysis systemsβ44May 20, 2022Updated 4 years ago
- Easy to use Audio Tagging in PyTorchβ23Aug 22, 2021Updated 4 years ago
- Reading list for research topics in Sound AIβ201Aug 8, 2024Updated last year
- Using YouTube to prepare a speech recognition dataset for any languageβ10Mar 30, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- β25Feb 26, 2026Updated 5 months ago
- β16Apr 11, 2019Updated 7 years ago
- Documentation of the Two!Ears Auditory Modelβ13Feb 14, 2019Updated 7 years ago
- A library of speech gadgets.β15Oct 15, 2022Updated 3 years ago
- Official Implementation of DMT: Dual Mean-Teacher in PyTorch.β10Oct 27, 2023Updated 2 years ago
- Permutation invariant training in PyTorchβ13Oct 2, 2020Updated 5 years ago
- Read audio with FFmpeg into NumPy/PyTorch via ctypes (standard library module)β11Aug 12, 2020Updated 5 years ago
- Detect emotion from audioβ14Nov 20, 2018Updated 7 years ago
- A recipe for creating a Speaker Identification system built on Kaldi.β15Jan 2, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Filter Bank Implementaion as Convolutional Neural Network using Python Kerasβ17Dec 18, 2024Updated last year
- Adapt Kaldi-ASR nnet3 chain models from Zamia-Speech.org to a different language modelβ33Jan 26, 2020Updated 6 years ago
- β22Jun 30, 2021Updated 5 years ago
- Evaluation script for VoxMovies dataset in PyTorchβ23Jan 12, 2024Updated 2 years ago
- My system for the DCASE 2022 Task 3 Sound Event Localizaiton and Detection.β12Nov 12, 2022Updated 3 years ago
- Development Toolkit for the VoxCeleb Speaker Recognition Challenge 2020β43Jul 17, 2020Updated 6 years ago
- Filtering and Noise Adding Toolβ29May 27, 2022Updated 4 years ago
- [ICLR 2022] "Audio Lottery: Speech Recognition Made Ultra-Lightweight, Noise-Robust, and Transferable", by Shaojin Ding, Tianlong Chen, Zβ¦β32Apr 8, 2022Updated 4 years ago
- Source code complementing our paper for acoustic event classification using convolutional neural networks.β70Jan 31, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A tool/script for batch speech data enhancement with speed/volume/RIRS/MUSANβ25Jun 28, 2020Updated 6 years ago
- Sound event detection with depthwise separable and dilated convolutions.β53Mar 30, 2020Updated 6 years ago
- A Diffrentiable WFST-based End-to-End Automatic Speech Recognition toollkit with flexible topology supportβ12Feb 15, 2026Updated 5 months ago
- The active learning algorithm, mismatch-first farthest-traversal. Implementation and visualization.β12Dec 25, 2021Updated 4 years ago
- In this repository, I try to combine k2 with speechbrain to decode well and fastly.β16Jun 17, 2022Updated 4 years ago
- π Easily apply audio-related machine learning models trained on the AudioSet dataset (527+ models/classes).β31Jun 17, 2024Updated 2 years ago
- Comprehensive Python library for speech and voice.β32Dec 8, 2022Updated 3 years ago
- β54Oct 17, 2023Updated 2 years ago
- Echo aware source separationβ13May 29, 2018Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Vim Speech Recognition Experimentsβ20May 30, 2025Updated last year
- Script to generate VAD dataset used in Asteroid recipeβ21Sep 30, 2021Updated 4 years ago
- Support for Clarity Enhancement and Prediction Challenges (obsolete - see README)β48Apr 14, 2022Updated 4 years ago
- Learning Domain-Invariant Transformation for Speaker Verification.β11Jun 13, 2023Updated 3 years ago
- Pypi installable TDNN and TDNN-F layers for PyTorch based acoustic model trainingβ41Dec 18, 2020Updated 5 years ago
- MirasVoice is a data set consisting speech samples from bilinguals to train neural network for optimization of speaker verification algorβ¦β19Mar 15, 2020Updated 6 years ago
- This repository describes our reproducible framework for assessing self-supervised representation learning from speechβ52Oct 8, 2021Updated 4 years ago