This repository is built with a focus on practical ways to obtain and work with the audio data of audioset. You can use this repository to download and precprocess audioset wav files for running the recipies of Audio Spectogram Transformer (AST) and Masked Autoencoder that listen (Audio - MAE).
☆18Jun 12, 2025Updated last year
Alternatives and similar repositories for Audioset
Users that are interested in Audioset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Subband system identification using generalized Weighted Overlap-Add (WOLA) filter bank for improved acoustic echo cancellation.☆17May 8, 2025Updated last year
- ☆20Aug 23, 2024Updated 2 years ago
- ☆46Apr 2, 2025Updated last year
- ☆27Jan 27, 2026Updated 8 months ago
- The official implementation of V-AURA: Temporally Aligned Audio for Video with Autoregression (ICASSP 2025) (Oral)☆36Feb 11, 2026Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆12Jun 17, 2017Updated 9 years ago
- Unofficial pytorch implementation of the paper "Learnable Fourier Features for Multi-Dimensional Spatial Positional Encoding", NeurIPS 20…☆13Apr 24, 2024Updated 2 years ago
- A PyTorch Dataset for Slakh2100☆10Feb 14, 2024Updated 2 years ago
- The code used to create the ARCA23K and ARCA23K-FSD datasets☆16Nov 9, 2021Updated 4 years ago
- [ICML'24] Creative Text-to-Audio Generation via Synthesizer Programming☆42Sep 26, 2024Updated 2 years ago
- SoloAudio: Target Sound Extraction with Language-oriented Audio Diffusion Transformer.☆124Jan 28, 2026Updated 8 months ago
- ☆53Mar 24, 2026Updated 6 months ago
- SongDriver2 achieves a balance between real-time emotion fit and soft transitions, enhancing the coherence of the generated music.☆111Nov 15, 2025Updated 10 months ago
- The repository of the paper: Wang et al., Learning interpretable representation for controllable polyphonic music generation, ISMIR 2020.☆45Mar 22, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆41May 12, 2025Updated last year
- Machine learning speaker characteristics☆46Updated this week
- Signal processing and deep learning approaches to detect partial discharge faults in covered conductors. An entry for a Kaggle competitio…☆17Mar 30, 2025Updated last year
- Modeling Harmonic Complexity using two models of Conditional Variational Autoencoders - MSc. Thesis☆10May 16, 2023Updated 3 years ago
- The electronic Holly Quran browser Elforkane☆11Nov 14, 2021Updated 4 years ago
- The implementation of "Instrument Separation of Symbolic Music by Explicitly Guided Diffusion Model"☆15Aug 16, 2022Updated 4 years ago
- ☆11May 4, 2020Updated 6 years ago
- JAVA and Docker based solution to host all service components (Owner, Manufacturer, Rendezvous) defined in SDO protocol. Reuses binaries …☆11Apr 5, 2023Updated 3 years ago
- code and demo of the ISMIR 2021 paper CollageNet☆12Jul 12, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- codebase for the Text-based NP Enrichment (TNE) paper☆19Mar 12, 2024Updated 2 years ago
- Implementation of "Improving Whispered Speech Recognition Performance using Pseudo-whispered based Data Augmentation"☆16Sep 11, 2026Updated 3 weeks ago
- Embedded Tajweed annotation for the Qur'an☆12Nov 30, 2025Updated 10 months ago
- Efficient Methods for BEamforming Deconvolution☆19Oct 26, 2017Updated 8 years ago
- Real Acoustic Fields An Audio-Visual Room Acoustics Dataset and Benchmark☆64Aug 29, 2024Updated 2 years ago
- code for A Large-scale Dataset for Audio-Language Representation Learning☆14Sep 18, 2024Updated 2 years ago
- code for "Automated and Intelligent Synthesis of Oxygen-Producing Catalysts from Martian Meteorites by Robotic AI-Chemist "☆12Jul 31, 2023Updated 3 years ago
- Learning an Interpretable End-to-End Network for Real-Time Acoustic Beamforming☆20Aug 20, 2024Updated 2 years ago
- ☆18May 15, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- S3PRL for Speech Emotion Recognition (see s3prl > downstream)☆15Feb 28, 2026Updated 7 months ago
- ISMIR 2021: Curriculum Learning for Imbalanced Classification in Large Vocabulary Automatic Chord Recognition☆10Nov 8, 2021Updated 4 years ago
- Zenodo Developers Site☆18Sep 30, 2026Updated last week
- ☆19Aug 27, 2018Updated 8 years ago
- ☆17Sep 2, 2017Updated 9 years ago
- ☆11Jan 22, 2017Updated 9 years ago
- D3.js visualization of the language parts of the 2011 Census of India.☆11Jul 19, 2026Updated 2 months ago