Examples for ICASSP2024 paper "StemGen: A music generation model that listens"
☆35Dec 19, 2023Updated 2 years ago
Alternatives and similar repositories for stemgen
Users that are interested in stemgen are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MiRA (Music Replication Assessment) tool is a model-independent open evaluation method based on four diverse audio music similarity metri…☆35Nov 14, 2025Updated 9 months ago
- Audio Prompt Adapter: Unleashing music editing abilities for text-to-music with lightweight finetuning [ISMIR 2024]☆57Nov 10, 2025Updated 9 months ago
- MelodyT5: A Unified Score-to-Score Transformer for Symbolic Music Processing [ISMIR 2024]☆50Jan 23, 2025Updated last year
- ☆12Feb 8, 2024Updated 2 years ago
- ☆32Jan 6, 2022Updated 4 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆35Sep 7, 2022Updated 4 years ago
- Codes and MIDI demos of ISMIR 2022 paper: Domain Adversarial Training on Conditional Variational Auto-Encoder for Controllable Music Gene…☆21Mar 28, 2023Updated 3 years ago
- Externalizer is a VST3 audio plugin developed using the JUCE framework, designed to enable routing audio to and from external hardware pr…☆15May 29, 2023Updated 3 years ago
- Enriching Music Descriptions with a Finetuned-LLM and Metadata for Text-to-Music Retrieval (TTMR++) [ICASSP24]☆42Oct 7, 2024Updated last year
- Official implementation of "Learning Music Audio Representations Via Weak Language Supervision" (ICASSP 2022)☆47Dec 3, 2024Updated last year
- This is the official implementation of MusER (AAAI'24).☆31Jun 4, 2025Updated last year
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into one☆26Aug 5, 2024Updated 2 years ago
- Supervised and unsupervised Concept-based explanation of pretrained music classifiers☆12Jul 27, 2023Updated 3 years ago
- The hybrid architecture is based on the idea that we could simply apply a GAN method (GANSpace) to another GAN model (GANSynth).☆25Aug 16, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for ICLR 2024 Paper: CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models☆23Jul 10, 2024Updated 2 years ago
- A Representation Evaluation Framework for Music Information Retrieval tasks☆56Apr 9, 2024Updated 2 years ago
- Codebase for 'A Real-Time Lyrics Alignment System Using Chroma And Phonetic Features For Classical Vocal Performance', ICASSP 2024☆14Oct 4, 2024Updated last year
- Pytorch implementation of automatic music transcription method that uses a two-level hierarchical frequency-time Transformer architecture…☆120Jul 11, 2023Updated 3 years ago
- The Song Describer dataset is an evaluation dataset made of ~1.1k captions for 706 permissively licensed music recordings.☆177Dec 22, 2023Updated 2 years ago
- MusAV: a dataset of relative arousal-valence annotations for validation of audio models☆17Dec 16, 2022Updated 3 years ago
- ScorePerformer: Expressive Piano Performance Rendering with Fine-Grained Control (ISMIR 2023)☆42Mar 10, 2025Updated last year
- Textless Speech-to-Music Retrieval Using Emotion Similarity [ICASSP23]☆17Aug 16, 2023Updated 3 years ago
- Codebase for the paper 'EncodecMAE: Leveraging neural codecs for universal audio representation learning'☆101Jul 24, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Baseline for DCASE 2024 Task 9: "Language-Queried Audio Source Separation"☆26Mar 27, 2024Updated 2 years ago
- ☆133Updated this week
- Fine-tune your own MusicGen with LoRA☆162Apr 26, 2024Updated 2 years ago
- Fast C++ implementation of ESOLA using KFRLib, can be used for online time-stretch augmentation during SpeechToText training.☆16Jul 25, 2020Updated 6 years ago
- Joint Embedding Predictive Architecture for Musical Stem Compatibility Estimation☆56Aug 6, 2024Updated 2 years ago
- A simple library for Fréchet Audio Distance (FAD) calculation☆266Aug 22, 2025Updated last year
- Compute distribution-based quality metrics for audio data using embeddings, with a focus on music.☆46Jan 15, 2026Updated 7 months ago
- The official implementation of TokenSynth (ICASSP 2025)☆95Jun 24, 2026Updated 2 months ago
- Repository for the paper "Combining audio control and style transfer using latent diffusion", accepted at ISMIR 2024☆69Feb 19, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Multitrack music mixing style transfer given a reference song using differentiable mixing console.☆65Jul 7, 2025Updated last year
- ☆38Jun 16, 2024Updated 2 years ago
- Pre-training, fine-tuning, and inference code with the MAEST models for music analysis applications.☆71Jun 27, 2025Updated last year
- Prosody and Pronunciation Modification Network☆64May 5, 2025Updated last year
- Convert meshes into physical audio models and play them by striking mesh vertices in a 3D viewer.☆18May 7, 2024Updated 2 years ago
- ZIQI-Eval: A Music Evaluation Benchmark for Large Language Models☆19Jul 23, 2024Updated 2 years ago
- ☆22Jul 27, 2026Updated last month