Solos: A Dataset for Audio-Visual Music Analysis
☆24Feb 17, 2023Updated 3 years ago
Alternatives and similar repositories for Solos
Users that are interested in Solos are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- code for "When Counterpoints Meet Chinese Folk Melody"☆11Feb 19, 2021Updated 5 years ago
- Backpropagable pytorch implementation of https://craffel.github.io/mir_eval/.☆35Jul 8, 2024Updated 2 years ago
- ☆12Apr 30, 2025Updated last year
- This repository holds datasets of polyphonic drum patterns used in the creation of Electronic Dance Music.☆16Dec 19, 2016Updated 9 years ago
- Code for the paper: Unified Gradient Reweighting for Model Biasing with Applications to Source Separation☆14Nov 16, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Unofficial implementation of music separation model by Luo et.al.☆13Nov 3, 2019Updated 6 years ago
- Project for MIDI to Audio Synthesis☆28Mar 13, 2023Updated 3 years ago
- This branch of Asteroid contains code for the vocal harmony and chamber ensemble separation related papers.☆12Nov 7, 2024Updated last year
- Music Demixing Challenge Submission Repo☆16Sep 8, 2023Updated 2 years ago
- ☆15May 10, 2026Updated 2 months ago
- ☆16Sep 7, 2022Updated 3 years ago
- Official implementation of A cappella: Audio-visual Singing VoiceSeparation, from BMVC21☆18May 14, 2022Updated 4 years ago
- ETH Zürich MSc Thesis: Accelerating Neural Audio Synthesis☆26Apr 10, 2023Updated 3 years ago
- ☆15Sep 24, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Baseline for DCASE 2024 Task 9: "Language-Queried Audio Source Separation"☆26Mar 27, 2024Updated 2 years ago
- STOI loss functions in PyTorch (mirror of https://github.com/mpariente/pytorch_stoi)☆15Aug 6, 2020Updated 5 years ago
- VoViT: Low Latency Graph-based Audio-Visual VoiceSeparation Transformer☆35Mar 18, 2023Updated 3 years ago
- ☆13May 9, 2022Updated 4 years ago
- Room acoustic simulator with a SOFA file loader.☆25Sep 27, 2024Updated last year
- The official implementation of V-AURA: Temporally Aligned Audio for Video with Autoregression (ICASSP 2025) (Oral)☆35Feb 11, 2026Updated 5 months ago
- Evaluation of a number of loudness meter implementations☆13Aug 28, 2021Updated 4 years ago
- An open agentic system built on smolagents, integrating multimodal state-of-the-art music AI models for understanding, generation, and in…☆32Feb 6, 2026Updated 5 months ago
- Epoch-synchronous overlap-add (ESOLA) for time-and pitch-scale modification of speech signals.☆23Jul 24, 2020Updated 6 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Streaming source separation for music and speech files, using the Open-Unmix LSTM architecture.☆21Dec 8, 2022Updated 3 years ago
- Code for Discriminative Sounding Objects Localization (NeurIPS 2020)☆61Jan 19, 2022Updated 4 years ago
- easy-to-use implementation of the ISMIR 2013 Audio Degradation Toolbox☆54Nov 19, 2019Updated 6 years ago
- [CVPR25] Official Implementation of CAV-MAE Sync☆31Apr 5, 2026Updated 3 months ago
- ☆49Jul 10, 2024Updated 2 years ago
- Zounds is a dataflow library for building directed acyclic graphs that transform audio. It uses the featureflow library to define the pro…☆24Dec 8, 2022Updated 3 years ago
- Source Separation training codebase for the Sound Demixing Challenge 2023.☆45May 18, 2023Updated 3 years ago
- Echo aware source separation☆13May 29, 2018Updated 8 years ago
- Crowdsourced Audio Quality Evaluation Toolkit☆55Dec 7, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆31Feb 4, 2021Updated 5 years ago
- An implementation of the Prism layer (https://arxiv.org/abs/2011.04823)☆12Nov 13, 2020Updated 5 years ago
- A Dataset for Cover Song Identification and Understanding☆66Feb 23, 2023Updated 3 years ago
- 🔊 A comprehensive list of open-source datasets for voice and sound computing (50+ datasets).☆20Apr 1, 2021Updated 5 years ago
- Utilities and experiments for training RAVE☆16Oct 23, 2023Updated 2 years ago
- Efficient Training for Multilingual Visual Speech Recognition: Pre-training with Discretized Visual Speech Representation (ACM MM 2024)☆20Mar 17, 2025Updated last year
- Experimenting with Lapped Transforms Jupyter Notebook☆14Jun 13, 2025Updated last year