[ECCV2022] D2M-GAN for music generation from dance videos
☆85Aug 16, 2022Updated 3 years ago
Alternatives and similar repositories for D2M-GAN
Users that are interested in D2M-GAN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML2023] Long-Term Rhythmic Video Soundtracker☆63Jul 28, 2025Updated last year
- [ICLR2023] Discrete Contrastive Diffusion for Cross-Modal Music and Image Generation (CDCD).☆163Apr 5, 2023Updated 3 years ago
- [ACM MM 2021 Best Paper Award] Video Background Music Generation with Controllable Music Transformer☆327Jun 8, 2025Updated last year
- NU-Wave 2: A General Neural Audio Upsampling Model for Various Sampling Rates [WIP]☆25Jul 5, 2022Updated 4 years ago
- ☆12Apr 30, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Python3 Implementation for 'Visual Rhythm and Beat' SIGGRAPH 2018☆20May 31, 2022Updated 4 years ago
- Code for CVPR 2022 paper "Bailando: 3D dance generation via Actor-Critic GPT with Choreographic Memory"☆435Dec 7, 2023Updated 2 years ago
- Parallel waveform generation with DiffusionGAN☆17Mar 26, 2022Updated 4 years ago
- Symphony Generation with Permutation Invariant Language Model☆256Oct 7, 2022Updated 3 years ago
- Official Implementation of "Multitrack Music Transformer" (ICASSP 2023)☆155Mar 14, 2024Updated 2 years ago
- Toward Universal Text-to-Music-Retrieval (TTMR) [ICASSP23]☆113Aug 12, 2023Updated 2 years ago
- PyTorch implementation of ECCV 2020 paper "Foley Music: Learning to Generate Music from Videos "☆39Dec 15, 2020Updated 5 years ago
- The source code of our paper "Diffsound: discrete diffusion model for text-to-sound generation"☆366Aug 3, 2023Updated 2 years ago
- Neural Lexicon Reader: Reduce Pronunciation Errors in End-to-end TTS by Leveraging External Textual Knowledge☆21Jul 25, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official PyTorch implementation of the paper "A Brand New Dance Partner:Music-Conditioned Pluralistic Dancing Synthesized by Multiple Dan…☆37Jul 6, 2022Updated 4 years ago
- PyTorch Implementation of Multi-Singer (ACM-MM'21)☆139May 8, 2022Updated 4 years ago
- Official implementation of DGP-based multi-speaker speech synthesis with PyTorch☆24Mar 23, 2021Updated 5 years ago
- Synthesis of MIDI with DDSP (https://midi-ddsp.github.io/)☆338Nov 30, 2022Updated 3 years ago
- An opensource music processing toolkit☆321Jun 25, 2023Updated 3 years ago
- A toolset for easy formant extraction and visualization from wav files and TTS models☆33Sep 2, 2022Updated 3 years ago
- DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code☆10Mar 8, 2022Updated 4 years ago
- Style-based Neural Drum Synthesis with GAN inversion☆33Nov 9, 2021Updated 4 years ago
- Chord-Conditioned Melody Harmonization with Controllable Harmonicity [ICASSP 2023]☆49Jul 15, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆24Mar 15, 2022Updated 4 years ago
- BRACE: The Breakdancing Competition Dataset for Dance Motion Synthesis☆79Jul 11, 2024Updated 2 years ago
- Evaluation metrics for machine-composed symbolic music. Paper: "The Jazz Transformer on the Front Line: Exploring the Shortcomings of AI-…☆64Oct 29, 2020Updated 5 years ago
- Adaptive Vocoder for Custom Voice☆61Sep 22, 2022Updated 3 years ago
- Autovocoder: Fast Waveform Generation from a Learned Speech Representation using Differentiable Digital Signal Processing☆71Dec 2, 2022Updated 3 years ago
- Making an AI-generated music video from any song with Wav2CLIP and VQGAN-CLIP☆245Jun 10, 2022Updated 4 years ago
- A Non-Autoregressive End-to-End Text-to-Speech (text-to-wav), supporting a family of SOTA unsupervised duration modelings. This project g…☆147Jun 6, 2022Updated 4 years ago
- A python package for high level musical data manipulation and preprocessing, making data ready to be fed to a neural network.☆43Jan 19, 2022Updated 4 years ago
- Source code for "FIGARO: Generating Symbolic Music with Fine-Grained Artistic Control"☆166Oct 15, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ICASSP 2022☆61Oct 12, 2021Updated 4 years ago
- API to support AIST++ Dataset: https://google.github.io/aistplusplus_dataset☆392Apr 10, 2023Updated 3 years ago
- ☆21Nov 29, 2022Updated 3 years ago
- ☆25Mar 12, 2022Updated 4 years ago
- Official implementation of SawSing (ISMIR'22)☆275Aug 28, 2022Updated 3 years ago
- Tegridy MIDI Dataset for precise and effective Music AI models creation.☆286Updated this week
- ☆10Apr 8, 2024Updated 2 years ago