This repository is built with a focus on practical ways to obtain and work with the audio data of audioset. You can use this repository to download and precprocess audioset wav files for running the recipies of Audio Spectogram Transformer (AST) and Masked Autoencoder that listen (Audio - MAE).
☆18Jun 12, 2025Updated last year
Alternatives and similar repositories for Audioset
Users that are interested in Audioset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Subband system identification using generalized Weighted Overlap-Add (WOLA) filter bank for improved acoustic echo cancellation.☆17May 8, 2025Updated last year
- ☆46Apr 2, 2025Updated last year
- ☆27Jan 27, 2026Updated 7 months ago
- The official implementation of V-AURA: Temporally Aligned Audio for Video with Autoregression (ICASSP 2025) (Oral)☆36Feb 11, 2026Updated 7 months ago
- ☆12Jun 17, 2017Updated 9 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Unofficial pytorch implementation of the paper "Learnable Fourier Features for Multi-Dimensional Spatial Positional Encoding", NeurIPS 20…☆13Apr 24, 2024Updated 2 years ago
- A PyTorch Dataset for Slakh2100☆10Feb 14, 2024Updated 2 years ago
- [ICML'24] Creative Text-to-Audio Generation via Synthesizer Programming☆42Sep 26, 2024Updated last year
- Free and Open Platform for AI-assisted Computing☆10May 19, 2019Updated 7 years ago
- SoloAudio: Target Sound Extraction with Language-oriented Audio Diffusion Transformer.☆124Jan 28, 2026Updated 7 months ago
- ☆53Mar 24, 2026Updated 5 months ago
- SongDriver2 achieves a balance between real-time emotion fit and soft transitions, enhancing the coherence of the generated music.☆111Nov 15, 2025Updated 10 months ago
- The repository of the paper: Wang et al., Learning interpretable representation for controllable polyphonic music generation, ISMIR 2020.☆45Mar 22, 2024Updated 2 years ago
- ☆41May 12, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [CVPR 2025] Pytorch implementation of the paper "Learning to Highlight Audio by Watching Movies"☆15Oct 1, 2025Updated 11 months ago
- Machine learning speaker characteristics☆46Updated this week
- Signal processing and deep learning approaches to detect partial discharge faults in covered conductors. An entry for a Kaggle competitio…☆17Mar 30, 2025Updated last year
- Modeling Harmonic Complexity using two models of Conditional Variational Autoencoders - MSc. Thesis☆10May 16, 2023Updated 3 years ago
- The electronic Holly Quran browser Elforkane☆11Nov 14, 2021Updated 4 years ago
- The implementation of "Instrument Separation of Symbolic Music by Explicitly Guided Diffusion Model"☆15Aug 16, 2022Updated 4 years ago
- JAVA and Docker based solution to host all service components (Owner, Manufacturer, Rendezvous) defined in SDO protocol. Reuses binaries …☆11Apr 5, 2023Updated 3 years ago
- code and demo of the ISMIR 2021 paper CollageNet☆12Jul 12, 2021Updated 5 years ago
- codebase for the Text-based NP Enrichment (TNE) paper☆19Mar 12, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of "Improving Whispered Speech Recognition Performance using Pseudo-whispered based Data Augmentation"☆15Sep 11, 2026Updated last week
- 使用Sentencepiece对中文语料进行分词☆13Nov 30, 2023Updated 2 years ago
- Embedded Tajweed annotation for the Qur'an☆12Nov 30, 2025Updated 9 months ago
- A Fast Hadamard Transform for Signals with Sub-linear Sparsity in the Transform Domain -- Implementation☆18Oct 10, 2016Updated 9 years ago
- Efficient Methods for BEamforming Deconvolution☆19Oct 26, 2017Updated 8 years ago
- Real Acoustic Fields An Audio-Visual Room Acoustics Dataset and Benchmark☆64Aug 29, 2024Updated 2 years ago
- code for A Large-scale Dataset for Audio-Language Representation Learning☆14Sep 18, 2024Updated 2 years ago
- Learning an Interpretable End-to-End Network for Real-Time Acoustic Beamforming☆20Aug 20, 2024Updated 2 years ago
- repository for rockchip rv1126B☆23Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆18May 15, 2021Updated 5 years ago
- S3PRL for Speech Emotion Recognition (see s3prl > downstream)☆15Feb 28, 2026Updated 6 months ago
- ISMIR 2021: Curriculum Learning for Imbalanced Classification in Large Vocabulary Automatic Chord Recognition☆10Nov 8, 2021Updated 4 years ago
- Zenodo Developers Site☆18Apr 10, 2026Updated 5 months ago
- ☆19Aug 27, 2018Updated 8 years ago
- ☆17Sep 2, 2017Updated 9 years ago
- ☆11Jan 22, 2017Updated 9 years ago