Examples for ICASSP2024 paper "StemGen: A music generation model that listens"
☆35Dec 19, 2023Updated 2 years ago
Alternatives and similar repositories for stemgen
Users that are interested in stemgen are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MiRA (Music Replication Assessment) tool is a model-independent open evaluation method based on four diverse audio music similarity metri…☆35Nov 14, 2025Updated 8 months ago
- Audio Prompt Adapter: Unleashing music editing abilities for text-to-music with lightweight finetuning [ISMIR 2024]☆57Nov 10, 2025Updated 8 months ago
- MelodyT5: A Unified Score-to-Score Transformer for Symbolic Music Processing [ISMIR 2024]☆50Jan 23, 2025Updated last year
- ☆11Feb 8, 2024Updated 2 years ago
- ☆32Jan 6, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆35Sep 7, 2022Updated 3 years ago
- Codes and MIDI demos of ISMIR 2022 paper: Domain Adversarial Training on Conditional Variational Auto-Encoder for Controllable Music Gene…☆21Mar 28, 2023Updated 3 years ago
- Enriching Music Descriptions with a Finetuned-LLM and Metadata for Text-to-Music Retrieval (TTMR++) [ICASSP24]☆43Oct 7, 2024Updated last year
- Official implementation of "Learning Music Audio Representations Via Weak Language Supervision" (ICASSP 2022)☆47Dec 3, 2024Updated last year
- This is the official implementation of MusER (AAAI'24).☆31Jun 4, 2025Updated last year
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into one☆26Aug 5, 2024Updated last year
- Supervised and unsupervised Concept-based explanation of pretrained music classifiers☆12Jul 27, 2023Updated 3 years ago
- The hybrid architecture is based on the idea that we could simply apply a GAN method (GANSpace) to another GAN model (GANSynth).☆25Aug 16, 2021Updated 4 years ago
- Code for ICLR 2024 Paper: CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models☆23Jul 10, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A Representation Evaluation Framework for Music Information Retrieval tasks☆54Apr 9, 2024Updated 2 years ago
- Codebase for 'A Real-Time Lyrics Alignment System Using Chroma And Phonetic Features For Classical Vocal Performance', ICASSP 2024☆14Oct 4, 2024Updated last year
- Pytorch implementation of automatic music transcription method that uses a two-level hierarchical frequency-time Transformer architecture…☆120Jul 11, 2023Updated 3 years ago
- The Song Describer dataset is an evaluation dataset made of ~1.1k captions for 706 permissively licensed music recordings.☆175Dec 22, 2023Updated 2 years ago
- MusAV: a dataset of relative arousal-valence annotations for validation of audio models☆17Dec 16, 2022Updated 3 years ago
- ScorePerformer: Expressive Piano Performance Rendering with Fine-Grained Control (ISMIR 2023)☆42Mar 10, 2025Updated last year
- Textless Speech-to-Music Retrieval Using Emotion Similarity [ICASSP23]☆17Aug 16, 2023Updated 2 years ago
- Codebase for the paper 'EncodecMAE: Leveraging neural codecs for universal audio representation learning'☆101Jul 24, 2024Updated 2 years ago
- Baseline for DCASE 2024 Task 9: "Language-Queried Audio Source Separation"☆26Mar 27, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆130Updated this week
- Fine-tune your own MusicGen with LoRA☆162Apr 26, 2024Updated 2 years ago
- Fast C++ implementation of ESOLA using KFRLib, can be used for online time-stretch augmentation during SpeechToText training.☆16Jul 25, 2020Updated 6 years ago
- Joint Embedding Predictive Architecture for Musical Stem Compatibility Estimation☆55Aug 6, 2024Updated last year
- A simple library for Fréchet Audio Distance (FAD) calculation☆266Aug 22, 2025Updated 11 months ago
- Compute distribution-based quality metrics for audio data using embeddings, with a focus on music.☆47Jan 15, 2026Updated 6 months ago
- The official implementation of TokenSynth (ICASSP 2025)☆92Jun 24, 2026Updated last month
- Repository for the paper "Combining audio control and style transfer using latent diffusion", accepted at ISMIR 2024☆67Feb 19, 2025Updated last year
- Multitrack music mixing style transfer given a reference song using differentiable mixing console.☆64Jul 7, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆38Jun 16, 2024Updated 2 years ago
- ☆20Updated this week
- Pre-training, fine-tuning, and inference code with the MAEST models for music analysis applications.☆70Jun 27, 2025Updated last year
- Prosody and Pronunciation Modification Network☆64May 5, 2025Updated last year
- Convert meshes into physical audio models and play them by striking mesh vertices in a 3D viewer.☆18May 7, 2024Updated 2 years ago
- ZIQI-Eval: A Music Evaluation Benchmark for Large Language Models☆18Jul 23, 2024Updated 2 years ago
- ☆107Oct 16, 2023Updated 2 years ago