Official implementation of the paper MGE-LDM: Joint Latent Diffusion for Simultaneous Music Generation and Source Extraction
☆20Feb 19, 2026Updated 6 months ago
Alternatives and similar repositories for MGE-LDM
Users that are interested in MGE-LDM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official repository of Myna: Masking-Based Contrastive Learning of Musical Representations☆17Mar 31, 2025Updated last year
- Code for the paper “Automatic Music Sample Identification with Multi-Track Contrastive Learning”.☆27May 22, 2026Updated 3 months ago
- ☆22Jan 3, 2026Updated 7 months ago
- ☆41Aug 9, 2026Updated 2 weeks ago
- [ACM MM 2025] AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation☆24Oct 28, 2025Updated 10 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆15Apr 13, 2025Updated last year
- ☆15Oct 13, 2025Updated 10 months ago
- ☆133Jul 23, 2026Updated last month
- JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment☆168Aug 7, 2025Updated last year
- A standardized toolkit of Kernel Audio Distance (KAD)—a distribution-free, unbiased, and computationally efficient metric for evaluating …☆105Jun 12, 2025Updated last year
- Official code for SongEcho☆66Mar 3, 2026Updated 5 months ago
- VST3 & AU plugin management for the https://github.com/spotify/pedalboard Python library☆15Jul 5, 2026Updated last month
- Encode and decode audio samples to/from continuous and discrete compressed representations!☆121Nov 25, 2025Updated 9 months ago
- Spot-if-AI helps users identify if a Spotify track has been generated with tools such as Suno or Udio.☆27Mar 25, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Pytorch project accompanying the paper "Comparing Deep Models and Evaluation Strategies for Multi-Pitch Estimation in Music Recordings", …☆15Aug 26, 2022Updated 4 years ago
- Audio Prompt Adapter: Unleashing music editing abilities for text-to-music with lightweight finetuning [ISMIR 2024]☆57Nov 10, 2025Updated 9 months ago
- Codebase for the paper 'EncodecMAE: Leveraging neural codecs for universal audio representation learning'☆101Jul 24, 2024Updated 2 years ago
- Enriching Music Descriptions with a Finetuned-LLM and Metadata for Text-to-Music Retrieval (TTMR++) [ICASSP24]☆42Oct 7, 2024Updated last year
- [ICLR 2026] Official implementation of Toward Complex-Valued Neural Networks for Waveform Generation☆21Apr 10, 2026Updated 4 months ago
- Official code and pretrained models for Linear Consistency Autoencoders (Lin-CAE), a method to induce linearity in audio autoencoders via…☆17Feb 12, 2026Updated 6 months ago
- ☆50May 22, 2026Updated 3 months ago
- ☆272Feb 14, 2024Updated 2 years ago
- ☆18Jan 20, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- LAFMA: A Latent Flow Matching Model for Text-to-Audio Generation (INTERSPEECH 2024)☆44Jun 13, 2024Updated 2 years ago
- This branch of Asteroid contains code for the vocal harmony and chamber ensemble separation related papers.☆12Nov 7, 2024Updated last year
- ☆32Jul 31, 2025Updated last year
- " Music Style Transfer with Time-Varying Inversion of Diffusion Models"☆59Jul 23, 2024Updated 2 years ago
- Generating drum loops using the Wave-U-Net conditioned on intuitive parameters.☆24Nov 19, 2020Updated 5 years ago
- [INTERSPEECH 2026] Pre-training, SFT, DPO and GRPO for Text-to-Audio Generation☆50Apr 17, 2026Updated 4 months ago
- [ICASSP2025] Official code for VoiceDiT: Dual-Condition Diffusion Transformer for Environment-Aware Speech Synthesis☆52Apr 9, 2025Updated last year
- MultiModal Audio Generation in Raw Waveform Space.☆155May 26, 2026Updated 3 months ago
- ☆12Dec 19, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The implementation of Paper: Compose Yourself: Average-Velocity Flow Matching for One-Step Speech Enhancement.☆21Aug 1, 2026Updated 3 weeks ago
- Official repo for BLAP: Bootstrapping Language-Audio Pre-training for Music Captioning presented at ICASSP 2025☆16Nov 18, 2024Updated last year
- Open, royalty free, lyrics2song / song generation data collection / cleaning pipeline.☆17May 9, 2025Updated last year
- A template for bootstrapping AbletonLive ControlSurface projects☆14Jan 31, 2022Updated 4 years ago
- [ACL 2026 Main] MeanAudio: Fast and Faithful Text-to-Audio Generation with Mean Flows☆150Sep 2, 2025Updated 11 months ago
- open-source Mandarian biased word dataset☆14Sep 21, 2023Updated 2 years ago
- [NeurIPS 2025] Separate Anything in Audio with Zero Training☆63Nov 3, 2025Updated 9 months ago