official code for CVPR'24 paper Diff-BGM
☆71Oct 12, 2024Updated last year
Alternatives and similar repositories for Diff-BGM
Users that are interested in Diff-BGM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2023] Video Background Music Generation: Dataset, Method and Evaluation☆78Mar 29, 2024Updated 2 years ago
- [CVPR 2025] Repository of VidMuse☆140Jun 7, 2025Updated last year
- ☆58Oct 10, 2024Updated last year
- Official implementation of Mozart's Touch: A Lightweight Multi-modal Music Generation Framework Based on Pre-Trained Large Models☆43Mar 17, 2026Updated 4 months ago
- Video Background Music Generation Using Unpaired Audio-Visual Data☆33Oct 8, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆32Nov 10, 2025Updated 8 months ago
- Mustango: Toward Controllable Text-to-Music Generation☆394Jun 2, 2025Updated last year
- Multimodal Music Generation with Explicit Bridges and Retrieval Augmentation: A framework for generating multimodal music by bridging dif…☆28Jan 21, 2025Updated last year
- Official source codes of airsep☆39Mar 26, 2024Updated 2 years ago
- Chord-Conditioned Melody Harmonization with Controllable Harmonicity [ICASSP 2023]☆49Jul 15, 2023Updated 3 years ago
- Polyffusion: A Diffusion Model for Polyphonic Score Generation with Internal and External Controls☆89Jul 16, 2024Updated 2 years ago
- Art2Mus is a system that generates music based on digitized artworks and text by using the AudioLDM2 architecture with an added projectio…☆20Oct 20, 2025Updated 9 months ago
- To make music production easier, we introduce Amadeus , a novel MIDI generation framework. While significantly improving generation quali…☆16Aug 29, 2025Updated 10 months ago
- Diff-Foley: Synchronized Video-to-Audio Synthesis with Latent Diffusion Models☆205May 29, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The official implementation of the IJCAI 2024 paper "MusicMagus: Zero-Shot Text-to-Music Editing via Diffusion Models".☆49Sep 11, 2024Updated last year
- [ACM MM 2021 Best Paper Award] Video Background Music Generation with Controllable Music Transformer☆326Jun 8, 2025Updated last year
- Official codes and models of the paper "Auffusion: Leveraging the Power of Diffusion and Large Language Models for Text-to-Audio Generati…☆194Mar 25, 2024Updated 2 years ago
- ☆88Oct 20, 2024Updated last year
- A library for computing Frechet Music Distance.☆31Feb 4, 2025Updated last year
- PyTorch Implementation of [AudioLCM]: a efficient and high-quality text-to-audio generation with latent consistency model.☆13Jun 15, 2024Updated 2 years ago
- Make-An-Audio-3: Transforming Text/Video into Audio via Flow-based Large Diffusion Transformers☆121May 19, 2025Updated last year
- Fine-tune your own MusicGen with LoRA☆161Apr 26, 2024Updated 2 years ago
- Official source codes of coco-mulla☆36Mar 21, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- MuChoMusic is a benchmark for evaluating music understanding in multimodal audio-language models.☆46Dec 3, 2024Updated last year
- The latent diffusion model for text-to-music generation.☆187Jan 26, 2024Updated 2 years ago
- This is the official repository for M2UGen☆513Jan 2, 2025Updated last year
- MelodyT5: A Unified Score-to-Score Transformer for Symbolic Music Processing [ISMIR 2024]☆50Jan 23, 2025Updated last year
- Fine-tune Stable Audio Open with DiT ControlNet.☆256May 16, 2025Updated last year
- This paper has been accepted in ACM ICMR 2021.☆20Nov 17, 2025Updated 8 months ago
- [ICML2023] Long-Term Rhythmic Video Soundtracker☆63Jul 28, 2025Updated 11 months ago
- ☆18Jan 20, 2025Updated last year
- A toy-like Text-to-Speech for Chinese/Mandarin synthesize, inspired by Tacotron & FastSpeech2 & RefineGAN.☆15May 25, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆13Sep 1, 2023Updated 2 years ago
- JamendoMaxCaps is a large-scale dataset of 362,000 instrumental creative commons tracks☆53May 24, 2025Updated last year
- Evaluation metrics for machine-composed symbolic music. Paper: "The Jazz Transformer on the Front Line: Exploring the Shortcomings of AI-…☆64Oct 29, 2020Updated 5 years ago
- Implementation of Multi-Source Music Generation with Latent Diffusion.☆29Sep 12, 2024Updated last year
- MuseControlLite: Multifunctional Music Generation with Lightweight Conditioners [ICML 2025]☆68Jan 6, 2026Updated 6 months ago
- [PyTorch] Minimal codebase for MusicGen models☆63Jan 7, 2025Updated last year
- This is the official repository of Emotion-Driven Melody Harmonization via Melodic Variation and Functional Representation.☆12Sep 25, 2024Updated last year