This package aims at simplifying the download of the AudioSet dataset.
☆60Jul 17, 2025Updated last year
Alternatives and similar repositories for audioset-download
Users that are interested in audioset-download are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Download audioset data super fastly with youtube-dl, ffmpeg and python multiprocessing☆48Aug 1, 2024Updated last year
- This package aims at simplifying the download of the AudioCaps dataset.☆35Dec 1, 2023Updated 2 years ago
- Toolkit for downloading and processing Google's AudioSet dataset.☆180Aug 22, 2025Updated 11 months ago
- WildDESED: A LLM-Powered Dataset for Wild Domestic Environment Sound Event Detection☆18Nov 19, 2024Updated last year
- Simple voice activity detection (VAD) algorithm in Python☆15Aug 10, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The official repo for Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation☆65Jul 2, 2025Updated last year
- FastSAG: Towards Fast Non-Autoregressive Singing Accompaniment Generation☆30Dec 19, 2024Updated last year
- [ICASSP2025] Official code for VoiceDiT: Dual-Condition Diffusion Transformer for Environment-Aware Speech Synthesis☆52Apr 9, 2025Updated last year
- Accompanying code for paper "Attention-Based Contextual Language Model Adaptation for Speech Recognition", submitted to ACL 2021.☆14Jul 25, 2023Updated 3 years ago
- provide SPHERE-formatted output as well as RIFF, AU, AIFF and raw☆14Dec 18, 2021Updated 4 years ago
- [ICASSP'24] Investigating Personalization Methods in Text to Music Generation☆47Mar 27, 2024Updated 2 years ago
- AudioLDM training, finetuning, evaluation and inference.☆304Dec 13, 2024Updated last year
- Code for ICASSP 2024 Paper: RECAP: Retrieval-Augmented Audio Captioning☆16Jun 23, 2024Updated 2 years ago
- A spoken version of the textual story cloze benchmark☆22Aug 6, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Pre-training BART model for the Italian Language☆16Dec 28, 2022Updated 3 years ago
- Understanding and Tackling Hallucinations in Large Audio-Language Models | ICASSP 2025, Interspeech 2024☆34Mar 14, 2025Updated last year
- A toolkit for benchmarking on a wide variety of audio deepfake datasets.☆36May 22, 2026Updated 2 months ago
- Unofficial implementation of FSD50k baselines for Sound Event Recognition☆27Apr 27, 2024Updated 2 years ago
- A 6-million Audio-Caption Paired Dataset Built with a LLMs and ALMs-based Automatic Pipeline☆208Dec 13, 2024Updated last year
- iSeparate library for the SDX2023 challenge☆15Dec 15, 2023Updated 2 years ago
- Code for paper Learning Audio-Visual Dereverberation☆32Aug 10, 2022Updated 3 years ago
- Code to train a custom time-domain autoencoder to dereverb audio☆16Nov 30, 2023Updated 2 years ago
- ☆10Jun 6, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- 🦇 Encoder of BAT (Learning to Reason about Spatial Sounds with Large Language Models)☆87Feb 13, 2025Updated last year
- Eureka-Audio: A 1.7B lightweight audio–language model that matches 7B–30B models on ASR, audio understanding, and paralinguistic reasonin…☆40Apr 11, 2026Updated 3 months ago
- SPEAR Challenge scripts and tools.☆25Mar 17, 2023Updated 3 years ago
- [INTERSPEECH 2024] Official pytorch code for the paper "Disentangled Representation Learning for Environment-agnostic Speaker Recognition…☆18Jul 23, 2024Updated 2 years ago
- ARCH: Audio Representations benCHmark☆57Aug 26, 2024Updated last year
- PyTorch implementation of the ICASSP-24 paper: "Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Superv…☆41Jan 6, 2024Updated 2 years ago
- ☆11Apr 12, 2024Updated 2 years ago
- Official repository for the paper "AudioMAE++: learning better masked audio representations with SwiGLU FFNs"☆15Apr 30, 2026Updated 3 months ago
- ☆39Jul 4, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Enhanced Reverberation As Supervision (ERAS) for unsupervised reverberant speech separation☆15Aug 1, 2024Updated last year
- ☆14Jun 16, 2023Updated 3 years ago
- The official code repo for "Zero-shot Audio Source Separation through Query-based Learning from Weakly-labeled Data", in AAAI 2022☆213Jul 14, 2022Updated 4 years ago
- This is the official train-dev-test release of the Interspeech2024 Discrete Speech Representation Challenge.☆32Jan 26, 2024Updated 2 years ago
- Generate accompaniment part with chords using Evolutionary algorithm.☆11May 8, 2022Updated 4 years ago
- Efficient Training for Multilingual Visual Speech Recognition: Pre-training with Discretized Visual Speech Representation (ACM MM 2024)☆20Mar 17, 2025Updated last year
- ☆23Feb 2, 2022Updated 4 years ago