☆33Nov 10, 2025Updated 10 months ago
Alternatives and similar repositories for GVMGen
Users that are interested in GVMGen are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Music production for silent film clips.☆34Apr 30, 2025Updated last year
- Source codes for the paper "Personalized Dynamic Music Emotion Recognition with Dual-Scale Attention-Based Meta-Learning" (PDMER) which p…☆14Mar 24, 2025Updated last year
- [ISMIR 2025] A curated list of vision-to-music generation: methods, datasets, evaluation and challenges.☆126Aug 9, 2025Updated last year
- Multimodal Music Generation with Explicit Bridges and Retrieval Augmentation: A framework for generating multimodal music by bridging dif…☆28Jan 21, 2025Updated last year
- The official code for “Dance-to-Music Generation with Encoder-based Textual Inversion“☆23Jun 17, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICCV 2023] Video Background Music Generation: Dataset, Method and Evaluation☆79Mar 29, 2024Updated 2 years ago
- Video Background Music Generation Using Unpaired Audio-Visual Data☆35Oct 8, 2024Updated last year
- [CVPR 2025] Repository of VidMuse☆150Jun 7, 2025Updated last year
- Datasets for affective music‑video retrieval☆13Aug 21, 2022Updated 4 years ago
- A library for computing Frechet Music Distance.☆32Feb 4, 2025Updated last year
- This is the repo with the code to conduct a comparative analysis of different audio representation models.☆11Aug 31, 2023Updated 3 years ago
- official code for CVPR'24 paper Diff-BGM☆71Oct 12, 2024Updated last year
- ☆59Oct 10, 2024Updated last year
- Video2Music: Suitable Music Generation from Videos using an Affective Multimodal Transformer model☆199Jul 30, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A curated list of Vision (video/image) to Audio Generation☆111Aug 24, 2026Updated last month
- Open, royalty free, lyrics2song / song generation data collection / cleaning pipeline.☆18May 9, 2025Updated last year
- A repo that builds text to music datasets from scratch, used in MuseContorlLite [ICML2025]☆28May 20, 2025Updated last year
- TunesFormer: Forming Irish Tunes with Control Codes by Bar Patching [HCMIR 2023]☆52Sep 19, 2023Updated 3 years ago
- This is the official implementation of RL-Chord (TNNLS).☆13Jan 2, 2024Updated 2 years ago
- This repository is for The Power of Sound(TPoS): Audio Reactive Video Generation with Stable Diffusion (ICCV2023)☆25Dec 7, 2023Updated 2 years ago
- This is the official implementation of MusER (AAAI'24).☆31Jun 4, 2025Updated last year
- Just a copy of https://github.com/RobynE23/CodeHS-Java-APCSA, but I added folders and some extra files that didn't exist. Another option …☆27Jan 23, 2024Updated 2 years ago
- Official implementation of Mozart's Touch: A Lightweight Multi-modal Music Generation Framework Based on Pre-Trained Large Models☆43Mar 17, 2026Updated 6 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- This is a cog implementation of the fine-tuner for Meta's MusicGen☆55Apr 5, 2024Updated 2 years ago
- Official Code Repository for the paper "Generating Realistic Images from In-the-wild Sounds", ICCV 2023☆12Aug 24, 2025Updated last year
- a python library for midi to wav, generation, visualization, which is design for machine learning☆11Mar 25, 2019Updated 7 years ago
- JamendoMaxCaps is a large-scale dataset of 362,000 instrumental creative commons tracks☆53May 24, 2025Updated last year
- Official repo for BLAP: Bootstrapping Language-Audio Pre-training for Music Captioning presented at ICASSP 2025☆16Nov 18, 2024Updated last year
- Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale☆29Aug 4, 2023Updated 3 years ago
- This tool can be used to re-project equidistant projected fisheye lens into an equirectangular projection image.☆11Sep 28, 2018Updated 7 years ago
- ☆14Nov 13, 2023Updated 2 years ago
- [ICML2023] Long-Term Rhythmic Video Soundtracker☆64Jul 28, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Official code for SongEcho☆67Mar 3, 2026Updated 6 months ago
- Metrics for evaluating music and audio generative models – with a focus on long-form, full-band, and stereo generations.☆304Sep 8, 2026Updated 2 weeks ago
- Neural Lexicon Reader: Reduce Pronunciation Errors in End-to-end TTS by Leveraging External Textual Knowledge☆21Jul 25, 2022Updated 4 years ago
- ☆43Apr 14, 2025Updated last year
- Synthesis of percussion sounds using sinusoidal modelling, DDSP noise synthesis, and a neural source filter approach.☆37Jan 7, 2025Updated last year
- Scaled diffusion transformer for text-to-speech synthesis (DiT + T5Gemma2 conditioning, TorchTitan & Megatron backends, tested up to 1024…☆25Mar 29, 2026Updated 5 months ago
- CLaMP 3: Universal Music Information Retrieval Across Unaligned Modalities and Unseen Languages [ACL 2025]☆257May 11, 2025Updated last year