This repository is built with a focus on practical ways to obtain and work with the audio data of audioset. You can use this repository to download and precprocess audioset wav files for running the recipies of Audio Spectogram Transformer (AST) and Masked Autoencoder that listen (Audio - MAE).
☆17Jun 12, 2025Updated last year
Alternatives and similar repositories for Audioset
Users that are interested in Audioset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Subband system identification using generalized Weighted Overlap-Add (WOLA) filter bank for improved acoustic echo cancellation.☆15May 8, 2025Updated last year
- ☆19Aug 23, 2024Updated last year
- ☆12Jun 17, 2017Updated 9 years ago
- [ICML'24] Creative Text-to-Audio Generation via Synthesizer Programming☆41Sep 26, 2024Updated last year
- ☆52Mar 24, 2026Updated 3 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- SongDriver2 achieves a balance between real-time emotion fit and soft transitions, enhancing the coherence of the generated music.☆11Nov 15, 2025Updated 8 months ago
- The repository of the paper: Wang et al., Learning interpretable representation for controllable polyphonic music generation, ISMIR 2020.☆45Mar 22, 2024Updated 2 years ago
- ☆40May 12, 2025Updated last year
- [CVPR 2025] Pytorch implementation of the paper "Learning to Highlight Audio by Watching Movies"☆15Oct 1, 2025Updated 9 months ago
- Signal processing and deep learning approaches to detect partial discharge faults in covered conductors. An entry for a Kaggle competitio…☆17Mar 30, 2025Updated last year
- Machine learning speaker characteristics☆46Jul 9, 2026Updated last week
- Modeling Harmonic Complexity using two models of Conditional Variational Autoencoders - MSc. Thesis☆10May 16, 2023Updated 3 years ago
- ☆11May 4, 2020Updated 6 years ago
- The electronic Holly Quran browser Elforkane☆11Nov 14, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implementation of "Improving Whispered Speech Recognition Performance using Pseudo-whispered based Data Augmentation"☆14Oct 31, 2024Updated last year
- code and demo of the ISMIR 2021 paper CollageNet☆12Jul 12, 2021Updated 5 years ago
- JAVA and Docker based solution to host all service components (Owner, Manufacturer, Rendezvous) defined in SDO protocol. Reuses binaries …☆11Apr 5, 2023Updated 3 years ago
- 使用Sentencepiece对中文语料进行分词☆13Nov 30, 2023Updated 2 years ago
- Efficient solution to the initialization selection of DNNs in transfer learning using duality diagrams as a similarity measure framework.☆10Nov 17, 2020Updated 5 years ago
- Learning an Interpretable End-to-End Network for Real-Time Acoustic Beamforming☆21Aug 20, 2024Updated last year
- Fourier Modal Method with Jax (official maintained repo)☆19Oct 21, 2025Updated 8 months ago
- A Fast Hadamard Transform for Signals with Sub-linear Sparsity in the Transform Domain -- Implementation☆18Oct 10, 2016Updated 9 years ago
- Embedded Tajweed annotation for the Qur'an☆11Nov 30, 2025Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- code for A Large-scale Dataset for Audio-Language Representation Learning☆14Sep 18, 2024Updated last year
- Official code for the paper "Provable Compositional Generalization for Object-Centric Learning" (ICLR 2024, oral)☆16Aug 26, 2024Updated last year
- ISMIR 2021: Curriculum Learning for Imbalanced Classification in Large Vocabulary Automatic Chord Recognition☆10Nov 8, 2021Updated 4 years ago
- SoloAudio: Target Sound Extraction with Language-oriented Audio Diffusion Transformer.☆119Jan 28, 2026Updated 5 months ago
- code for "Automated and Intelligent Synthesis of Oxygen-Producing Catalysts from Martian Meteorites by Robotic AI-Chemist "☆12Jul 31, 2023Updated 2 years ago
- ☆18May 15, 2021Updated 5 years ago
- Deploy ChatGLM on Modelz☆16Mar 20, 2023Updated 3 years ago
- S3PRL for Speech Emotion Recognition (see s3prl > downstream)☆15Feb 28, 2026Updated 4 months ago
- Zenodo Developers Site☆17Apr 10, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Performative Latents for Adaptive Unsupervised DDSP (PLAUD)☆20Jul 9, 2026Updated last week
- ☆11Jan 22, 2017Updated 9 years ago
- ☆17Sep 2, 2017Updated 8 years ago
- D3.js visualization of the language parts of the 2011 Census of India.☆11Updated this week
- Tools to isolate speaker and transcribe unstructured audio clips☆11Dec 4, 2022Updated 3 years ago
- This repo contains the code to reproduce the paper: "Enriched Music Representations with Multiple Cross-modal Contrastive Learning"☆15Jun 22, 2023Updated 3 years ago
- The official implementation of V-AURA: Temporally Aligned Audio for Video with Autoregression (ICASSP 2025) (Oral)☆35Feb 11, 2026Updated 5 months ago