1st place solution to the DCASE 2020 - Task 5 - Urban Sound Tagging with Spatiotemporal Context
☆17Dec 8, 2022Updated 3 years ago
Alternatives and similar repositories for urban-sound-tagging
Users that are interested in urban-sound-tagging are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Urban Sound Classification : striving towards a fair comparison☆17Dec 11, 2020Updated 5 years ago
- JAMS annotation files for the original and augmented UrbanSound8K dataset☆35Jan 31, 2018Updated 8 years ago
- Event Relation in Text-to-Audio (TTA) Generation☆21Feb 26, 2025Updated last year
- PyTorch Implementation of SubSpectralNet - Using Sub-Spectrogram based Convolutional Neural Networks for Acoustic Scene Classification, a…☆21Feb 20, 2019Updated 7 years ago
- SubSpectralNet - Using Sub-Spectrogram based Convolutional Neural Networks for Acoustic Scene Classification, accepted in ICASSP 2019☆18Feb 20, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 1st place solution to the DCASE 2019 - Task 5 - Urban Sound Tagging☆30Mar 19, 2021Updated 5 years ago
- ☆20May 13, 2019Updated 7 years ago
- CycleGAN-based Emotion Style Transfer as Data Augmentation for Speech Emotion Recognition☆12Oct 7, 2019Updated 6 years ago
- A pytorch implementation of the paper : Acoustic Scene Classification with Multiple Decision Schemes.☆20Dec 12, 2020Updated 5 years ago
- Code for the 3rd place solution to Freesound Audio Tagging 2019 Challenge☆54Dec 8, 2022Updated 3 years ago
- ☆10Feb 3, 2022Updated 4 years ago
- Aty-TTS: Improving fairness for spoken language understanding in atypical speech with Text-to-Speech☆12May 14, 2025Updated last year
- ☆14Dec 12, 2018Updated 7 years ago
- Documentation and programming examples☆12Jun 22, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Learning discriminative and robust time-frequency representations for environmental sound classification: Convolutional neural networks (…☆31Dec 19, 2019Updated 6 years ago
- MaskFlow: Discrete Flows For Flexible and Efficient Long Video Generation☆28Mar 4, 2025Updated last year
- Tacotron2 with BERT examples☆10Jul 8, 2019Updated 7 years ago
- radiomixer☆14Feb 16, 2022Updated 4 years ago
- 2nd place solution for 2020 DCASE challenge task 6 audio captioning. http://dcase.community/challenge2020/task-automatic-audio-captioning…☆24Aug 3, 2023Updated 3 years ago
- The python implementation for paper "Towards Discriminative Representation Learning for Speech Emotion Recognition" in IJCAI-2019☆23Aug 12, 2019Updated 7 years ago
- Thai smart home corpus with "Gowajee" hotword☆20Jul 30, 2023Updated 3 years ago
- System that ranks 2nd in DCASE 2022 Challenge Task 5: Few-shot Bioacoustic Event Detection☆28Jul 6, 2022Updated 4 years ago
- Baseline of DCASE 2020 task 4☆42Oct 24, 2022Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Histogram Layer Time Delay Neural Networks For Passive Sonar Classification☆19Jan 21, 2026Updated 6 months ago
- Podcast Summarizer with LLM Technology☆30May 28, 2025Updated last year
- A Dockerized Jupyter notebook environment with pre-installed audio machine learning tools.☆12Feb 28, 2019Updated 7 years ago
- ☆21Jul 15, 2024Updated 2 years ago
- Real-time end-to-end singing voice convertion☆25Nov 3, 2024Updated last year
- Repo associated to the DESED dataset, download and creation of data☆155Jul 16, 2024Updated 2 years ago
- AsoSoft Speech Corpus for Central-Kurdish Text-To-Speech☆23Jun 24, 2022Updated 4 years ago
- Deep learning and standard machine learning methods are developed and compared in classfying audio samples from microphones deployed abo…☆11Jan 17, 2020Updated 6 years ago
- CP-JKU submission to DCASE 20☆44Apr 19, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A transcribed speech dataset in Wolof, Pulaar and Sereer, to support agriculture. Funded by Lacuna Fund.☆21Mar 26, 2026Updated 4 months ago
- Acoustic echo cancelation(AEC) is a main algorithm in the pipe line of acoustic devices with KWS or ASR. FNLMS is used.☆19Apr 22, 2019Updated 7 years ago
- Learning Complex Basis Functions for Invariant Signal Representations with the Complex Autoencoder☆38Dec 16, 2024Updated last year
- ASR text preprocessing utility☆21Aug 5, 2024Updated 2 years ago
- some fast fourier wav file analysis scripts in python☆10Apr 25, 2018Updated 8 years ago
- ☆14Oct 2, 2017Updated 8 years ago
- Soundscape Ecology Toolkit☆11Mar 25, 2016Updated 10 years ago