This repository is built with a focus on practical ways to obtain and work with the audio data of audioset. You can use this repository to download and precprocess audioset wav files for running the recipies of Audio Spectogram Transformer (AST) and Masked Autoencoder that listen (Audio - MAE).
☆17Jun 12, 2025Updated last year
Alternatives and similar repositories for Audioset
Users that are interested in Audioset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Subband system identification using generalized Weighted Overlap-Add (WOLA) filter bank for improved acoustic echo cancellation.☆17May 8, 2025Updated last year
- ☆20Aug 23, 2024Updated 2 years ago
- ☆47Apr 2, 2025Updated last year
- The official implementation of V-AURA: Temporally Aligned Audio for Video with Autoregression (ICASSP 2025) (Oral)☆35Feb 11, 2026Updated 6 months ago
- ☆12Jun 17, 2017Updated 9 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A PyTorch Dataset for Slakh2100☆10Feb 14, 2024Updated 2 years ago
- The code used to create the ARCA23K and ARCA23K-FSD datasets☆16Nov 9, 2021Updated 4 years ago
- Free and Open Platform for AI-assisted Computing☆10May 19, 2019Updated 7 years ago
- ☆53Mar 24, 2026Updated 5 months ago
- SongDriver2 achieves a balance between real-time emotion fit and soft transitions, enhancing the coherence of the generated music.☆111Nov 15, 2025Updated 9 months ago
- The repository of the paper: Wang et al., Learning interpretable representation for controllable polyphonic music generation, ISMIR 2020.☆45Mar 22, 2024Updated 2 years ago
- ☆40May 12, 2025Updated last year
- [CVPR 2025] Pytorch implementation of the paper "Learning to Highlight Audio by Watching Movies"☆15Oct 1, 2025Updated 10 months ago
- Machine learning speaker characteristics☆46Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Signal processing and deep learning approaches to detect partial discharge faults in covered conductors. An entry for a Kaggle competitio…☆17Mar 30, 2025Updated last year
- Modeling Harmonic Complexity using two models of Conditional Variational Autoencoders - MSc. Thesis☆10May 16, 2023Updated 3 years ago
- The electronic Holly Quran browser Elforkane☆11Nov 14, 2021Updated 4 years ago
- The implementation of "Instrument Separation of Symbolic Music by Explicitly Guided Diffusion Model"☆15Aug 16, 2022Updated 4 years ago
- ☆11May 4, 2020Updated 6 years ago
- JAVA and Docker based solution to host all service components (Owner, Manufacturer, Rendezvous) defined in SDO protocol. Reuses binaries …☆11Apr 5, 2023Updated 3 years ago
- code and demo of the ISMIR 2021 paper CollageNet☆12Jul 12, 2021Updated 5 years ago
- codebase for the Text-based NP Enrichment (TNE) paper☆19Mar 12, 2024Updated 2 years ago
- Implementation of "Improving Whispered Speech Recognition Performance using Pseudo-whispered based Data Augmentation"☆15Oct 31, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Efficient solution to the initialization selection of DNNs in transfer learning using duality diagrams as a similarity measure framework.☆10Nov 17, 2020Updated 5 years ago
- Embedded Tajweed annotation for the Qur'an☆11Nov 30, 2025Updated 8 months ago
- 使用Sentencepiece对中文语料进行分词☆13Nov 30, 2023Updated 2 years ago
- Efficient Methods for BEamforming Deconvolution☆20Oct 26, 2017Updated 8 years ago
- code for A Large-scale Dataset for Audio-Language Representation Learning☆14Sep 18, 2024Updated last year
- code for "Automated and Intelligent Synthesis of Oxygen-Producing Catalysts from Martian Meteorites by Robotic AI-Chemist "☆12Jul 31, 2023Updated 3 years ago
- Learning an Interpretable End-to-End Network for Real-Time Acoustic Beamforming☆21Aug 20, 2024Updated 2 years ago
- ☆18May 15, 2021Updated 5 years ago
- S3PRL for Speech Emotion Recognition (see s3prl > downstream)☆15Feb 28, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ISMIR 2021: Curriculum Learning for Imbalanced Classification in Large Vocabulary Automatic Chord Recognition☆10Nov 8, 2021Updated 4 years ago
- Zenodo Developers Site☆18Apr 10, 2026Updated 4 months ago
- ☆19Aug 27, 2018Updated 8 years ago
- ☆17Sep 2, 2017Updated 8 years ago
- ☆11Jan 22, 2017Updated 9 years ago
- D3.js visualization of the language parts of the 2011 Census of India.☆11Jul 19, 2026Updated last month
- Tools to isolate speaker and transcribe unstructured audio clips☆11Dec 4, 2022Updated 3 years ago