Examples for ICASSP2024 paper "StemGen: A music generation model that listens"
☆35Dec 19, 2023Updated 2 years ago
Alternatives and similar repositories for stemgen
Users that are interested in stemgen are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MiRA (Music Replication Assessment) tool is a model-independent open evaluation method based on four diverse audio music similarity metri…☆35Nov 14, 2025Updated 10 months ago
- Audio Prompt Adapter: Unleashing music editing abilities for text-to-music with lightweight finetuning [ISMIR 2024]☆57Nov 10, 2025Updated 10 months ago
- MelodyT5: A Unified Score-to-Score Transformer for Symbolic Music Processing [ISMIR 2024]☆51Jan 23, 2025Updated last year
- ☆12Feb 8, 2024Updated 2 years ago
- ☆32Jan 6, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆35Sep 7, 2022Updated 4 years ago
- Codes and MIDI demos of ISMIR 2022 paper: Domain Adversarial Training on Conditional Variational Auto-Encoder for Controllable Music Gene…☆21Mar 28, 2023Updated 3 years ago
- Externalizer is a VST3 audio plugin developed using the JUCE framework, designed to enable routing audio to and from external hardware pr…☆15May 29, 2023Updated 3 years ago
- Enriching Music Descriptions with a Finetuned-LLM and Metadata for Text-to-Music Retrieval (TTMR++) [ICASSP24]☆42Oct 7, 2024Updated last year
- Official implementation of "Learning Music Audio Representations Via Weak Language Supervision" (ICASSP 2022)☆47Dec 3, 2024Updated last year
- This is the official implementation of MusER (AAAI'24).☆31Jun 4, 2025Updated last year
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into one☆26Aug 5, 2024Updated 2 years ago
- Supervised and unsupervised Concept-based explanation of pretrained music classifiers☆12Jul 27, 2023Updated 3 years ago
- The hybrid architecture is based on the idea that we could simply apply a GAN method (GANSpace) to another GAN model (GANSynth).☆25Aug 16, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code for ICLR 2024 Paper: CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models☆23Jul 10, 2024Updated 2 years ago
- A Representation Evaluation Framework for Music Information Retrieval tasks☆56Apr 9, 2024Updated 2 years ago
- Codebase for 'A Real-Time Lyrics Alignment System Using Chroma And Phonetic Features For Classical Vocal Performance', ICASSP 2024☆14Oct 4, 2024Updated last year
- Pytorch implementation of automatic music transcription method that uses a two-level hierarchical frequency-time Transformer architecture…☆120Jul 11, 2023Updated 3 years ago
- The Song Describer dataset is an evaluation dataset made of ~1.1k captions for 706 permissively licensed music recordings.☆178Dec 22, 2023Updated 2 years ago
- MusAV: a dataset of relative arousal-valence annotations for validation of audio models☆17Dec 16, 2022Updated 3 years ago
- ScorePerformer: Expressive Piano Performance Rendering with Fine-Grained Control (ISMIR 2023)☆42Mar 10, 2025Updated last year
- Textless Speech-to-Music Retrieval Using Emotion Similarity [ICASSP23]☆17Aug 16, 2023Updated 3 years ago
- Codebase for the paper 'EncodecMAE: Leveraging neural codecs for universal audio representation learning'☆101Jul 24, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Baseline for DCASE 2024 Task 9: "Language-Queried Audio Source Separation"☆26Mar 27, 2024Updated 2 years ago
- ☆134Sep 1, 2026Updated 3 weeks ago
- Fine-tune your own MusicGen with LoRA☆162Apr 26, 2024Updated 2 years ago
- Fast C++ implementation of ESOLA using KFRLib, can be used for online time-stretch augmentation during SpeechToText training.☆16Jul 25, 2020Updated 6 years ago
- Joint Embedding Predictive Architecture for Musical Stem Compatibility Estimation☆56Aug 6, 2024Updated 2 years ago
- A simple library for Fréchet Audio Distance (FAD) calculation☆266Aug 22, 2025Updated last year
- Compute distribution-based quality metrics for audio data using embeddings, with a focus on music.☆46Jan 15, 2026Updated 8 months ago
- The official implementation of TokenSynth (ICASSP 2025)☆97Jun 24, 2026Updated 3 months ago
- Repository for the paper "Combining audio control and style transfer using latent diffusion", accepted at ISMIR 2024☆70Feb 19, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Multitrack music mixing style transfer given a reference song using differentiable mixing console.☆65Jul 7, 2025Updated last year
- ☆38Jun 16, 2024Updated 2 years ago
- Pre-training, fine-tuning, and inference code with the MAEST models for music analysis applications.☆71Jun 27, 2025Updated last year
- Prosody and Pronunciation Modification Network☆64May 5, 2025Updated last year
- ZIQI-Eval: A Music Evaluation Benchmark for Large Language Models☆19Jul 23, 2024Updated 2 years ago
- Convert meshes into physical audio models and play them by striking mesh vertices in a 3D viewer.☆18May 7, 2024Updated 2 years ago
- ☆106Oct 16, 2023Updated 2 years ago