1st place solution to the DCASE 2020 - Task 5 - Urban Sound Tagging with Spatiotemporal Context
☆17Dec 8, 2022Updated 3 years ago
Alternatives and similar repositories for urban-sound-tagging
Users that are interested in urban-sound-tagging are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Urban Sound Classification : striving towards a fair comparison☆17Dec 11, 2020Updated 5 years ago
- JAMS annotation files for the original and augmented UrbanSound8K dataset☆35Jan 31, 2018Updated 8 years ago
- Event Relation in Text-to-Audio (TTA) Generation☆22Feb 26, 2025Updated last year
- PyTorch Implementation of SubSpectralNet - Using Sub-Spectrogram based Convolutional Neural Networks for Acoustic Scene Classification, a…☆21Feb 20, 2019Updated 7 years ago
- SubSpectralNet - Using Sub-Spectrogram based Convolutional Neural Networks for Acoustic Scene Classification, accepted in ICASSP 2019☆18Feb 20, 2019Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆20May 13, 2019Updated 7 years ago
- CycleGAN-based Emotion Style Transfer as Data Augmentation for Speech Emotion Recognition☆12Oct 7, 2019Updated 6 years ago
- A pytorch implementation of the paper : Acoustic Scene Classification with Multiple Decision Schemes.☆20Dec 12, 2020Updated 5 years ago
- Code for the 3rd place solution to Freesound Audio Tagging 2019 Challenge☆54Dec 8, 2022Updated 3 years ago
- Learning discriminative and robust time-frequency representations for environmental sound classification: Convolutional neural networks (…☆31Dec 19, 2019Updated 6 years ago
- 2021年全国水下机器人算法大赛-光学赛道☆15Jul 5, 2021Updated 5 years ago
- MaskFlow: Discrete Flows For Flexible and Efficient Long Video Generation☆28Mar 4, 2025Updated last year
- Tacotron2 with BERT examples☆10Jul 8, 2019Updated 7 years ago
- radiomixer☆14Feb 16, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 2nd place solution for 2020 DCASE challenge task 6 audio captioning. http://dcase.community/challenge2020/task-automatic-audio-captioning…☆24Aug 3, 2023Updated 3 years ago
- The python implementation for paper "Towards Discriminative Representation Learning for Speech Emotion Recognition" in IJCAI-2019☆23Aug 12, 2019Updated 7 years ago
- Thai smart home corpus with "Gowajee" hotword☆21Jul 30, 2023Updated 3 years ago
- System that ranks 2nd in DCASE 2022 Challenge Task 5: Few-shot Bioacoustic Event Detection☆28Jul 6, 2022Updated 4 years ago
- Histogram Layer Time Delay Neural Networks For Passive Sonar Classification☆19Jan 21, 2026Updated 7 months ago
- Podcast Summarizer with LLM Technology☆30May 28, 2025Updated last year
- A Dockerized Jupyter notebook environment with pre-installed audio machine learning tools.☆12Feb 28, 2019Updated 7 years ago
- ☆21Jul 15, 2024Updated 2 years ago
- ☆11Mar 15, 2017Updated 9 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Real-time end-to-end singing voice convertion☆25Nov 3, 2024Updated last year
- Repo associated to the DESED dataset, download and creation of data☆155Jul 16, 2024Updated 2 years ago
- AsoSoft Speech Corpus for Central-Kurdish Text-To-Speech☆23Jun 24, 2022Updated 4 years ago
- CP-JKU submission to DCASE 20☆44Apr 19, 2021Updated 5 years ago
- A transcribed speech dataset in Wolof, Pulaar and Sereer, to support agriculture. Funded by Lacuna Fund.☆22Mar 26, 2026Updated 5 months ago
- Acoustic echo cancelation(AEC) is a main algorithm in the pipe line of acoustic devices with KWS or ASR. FNLMS is used.☆19Apr 22, 2019Updated 7 years ago
- A Starlette example for deployment in fastai2☆11Dec 18, 2020Updated 5 years ago
- Simple script to re-rank images using OpenAI's CLIP https://github.com/openai/CLIP.☆15May 3, 2021Updated 5 years ago
- ASR text preprocessing utility☆21Aug 5, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- some fast fourier wav file analysis scripts in python☆10Apr 25, 2018Updated 8 years ago
- ☆14Oct 2, 2017Updated 8 years ago
- Soundscape Ecology Toolkit☆11Mar 25, 2016Updated 10 years ago
- ☆23Aug 11, 2020Updated 6 years ago
- Homemade LightGBM and VGG-net experiment setup for DCASE2017 task 1☆11Aug 8, 2017Updated 9 years ago
- ☆33Dec 23, 2025Updated 8 months ago
- Deep Audio Segmenter, unsupervised☆10Feb 20, 2026Updated 6 months ago