Examples for ICASSP2024 paper "StemGen: A music generation model that listens"
☆35Dec 19, 2023Updated 2 years ago
Alternatives and similar repositories for stemgen
Users that are interested in stemgen are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MiRA (Music Replication Assessment) tool is a model-independent open evaluation method based on four diverse audio music similarity metri…☆35Nov 14, 2025Updated 9 months ago
- Audio Prompt Adapter: Unleashing music editing abilities for text-to-music with lightweight finetuning [ISMIR 2024]☆57Nov 10, 2025Updated 9 months ago
- MelodyT5: A Unified Score-to-Score Transformer for Symbolic Music Processing [ISMIR 2024]☆50Jan 23, 2025Updated last year
- ☆12Feb 8, 2024Updated 2 years ago
- ☆32Jan 6, 2022Updated 4 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆35Sep 7, 2022Updated 3 years ago
- Codes and MIDI demos of ISMIR 2022 paper: Domain Adversarial Training on Conditional Variational Auto-Encoder for Controllable Music Gene…☆21Mar 28, 2023Updated 3 years ago
- Enriching Music Descriptions with a Finetuned-LLM and Metadata for Text-to-Music Retrieval (TTMR++) [ICASSP24]☆42Oct 7, 2024Updated last year
- Official implementation of "Learning Music Audio Representations Via Weak Language Supervision" (ICASSP 2022)☆47Dec 3, 2024Updated last year
- This is the official implementation of MusER (AAAI'24).☆31Jun 4, 2025Updated last year
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into one☆26Aug 5, 2024Updated 2 years ago
- Supervised and unsupervised Concept-based explanation of pretrained music classifiers☆12Jul 27, 2023Updated 3 years ago
- The hybrid architecture is based on the idea that we could simply apply a GAN method (GANSpace) to another GAN model (GANSynth).☆25Aug 16, 2021Updated 5 years ago
- Code for ICLR 2024 Paper: CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models☆23Jul 10, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A Representation Evaluation Framework for Music Information Retrieval tasks☆55Apr 9, 2024Updated 2 years ago
- Codebase for 'A Real-Time Lyrics Alignment System Using Chroma And Phonetic Features For Classical Vocal Performance', ICASSP 2024☆14Oct 4, 2024Updated last year
- Pytorch implementation of automatic music transcription method that uses a two-level hierarchical frequency-time Transformer architecture…☆120Jul 11, 2023Updated 3 years ago
- The Song Describer dataset is an evaluation dataset made of ~1.1k captions for 706 permissively licensed music recordings.☆176Dec 22, 2023Updated 2 years ago
- MusAV: a dataset of relative arousal-valence annotations for validation of audio models☆17Dec 16, 2022Updated 3 years ago
- ScorePerformer: Expressive Piano Performance Rendering with Fine-Grained Control (ISMIR 2023)☆42Mar 10, 2025Updated last year
- Textless Speech-to-Music Retrieval Using Emotion Similarity [ICASSP23]☆17Aug 16, 2023Updated 3 years ago
- Codebase for the paper 'EncodecMAE: Leveraging neural codecs for universal audio representation learning'☆101Jul 24, 2024Updated 2 years ago
- Baseline for DCASE 2024 Task 9: "Language-Queried Audio Source Separation"☆26Mar 27, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆132Jul 23, 2026Updated 3 weeks ago
- Fine-tune your own MusicGen with LoRA☆162Apr 26, 2024Updated 2 years ago
- Fast C++ implementation of ESOLA using KFRLib, can be used for online time-stretch augmentation during SpeechToText training.☆16Jul 25, 2020Updated 6 years ago
- Joint Embedding Predictive Architecture for Musical Stem Compatibility Estimation☆55Aug 6, 2024Updated 2 years ago
- A simple library for Fréchet Audio Distance (FAD) calculation☆266Aug 22, 2025Updated 11 months ago
- Compute distribution-based quality metrics for audio data using embeddings, with a focus on music.☆47Jan 15, 2026Updated 7 months ago
- The official implementation of TokenSynth (ICASSP 2025)☆94Jun 24, 2026Updated last month
- Repository for the paper "Combining audio control and style transfer using latent diffusion", accepted at ISMIR 2024☆67Feb 19, 2025Updated last year
- Multitrack music mixing style transfer given a reference song using differentiable mixing console.☆65Jul 7, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆38Jun 16, 2024Updated 2 years ago
- ☆21Jul 27, 2026Updated 3 weeks ago
- Pre-training, fine-tuning, and inference code with the MAEST models for music analysis applications.☆71Jun 27, 2025Updated last year
- Prosody and Pronunciation Modification Network☆64May 5, 2025Updated last year
- Convert meshes into physical audio models and play them by striking mesh vertices in a 3D viewer.☆18May 7, 2024Updated 2 years ago
- ZIQI-Eval: A Music Evaluation Benchmark for Large Language Models☆18Jul 23, 2024Updated 2 years ago
- ☆107Oct 16, 2023Updated 2 years ago