☆22Jun 18, 2026Updated last month
Alternatives and similar repositories for VocalParse
Users that are interested in VocalParse are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Oct 13, 2025Updated 9 months ago
- code and demo of the ISMIR 2021 paper CollageNet☆12Jul 12, 2021Updated 5 years ago
- Official code for SongEcho☆64Mar 3, 2026Updated 5 months ago
- ☆19Jan 19, 2026Updated 6 months ago
- Containing SOTA methods that follows time-varying conditions for Text-to-Music☆24Jan 1, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ACM MM 2025] AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation☆24Oct 28, 2025Updated 9 months ago
- An All-in-One Speech, Sound, Music Codec with Single Nested Codebook☆28Oct 11, 2025Updated 9 months ago
- YingMusic-Singer-Plus: Controllable Singing Voice Synthesis with Flexible Lyric Manipulation and Annotation-free Melody Guidance☆86Apr 12, 2026Updated 3 months ago
- The official implementation of "TONet: Tone-Octave Network for Singing Melody Extraction from Polyphonic Music"☆42Oct 25, 2022Updated 3 years ago
- Official code and pretrained models for Linear Consistency Autoencoders (Lin-CAE), a method to induce linearity in audio autoencoders via…☆17Feb 12, 2026Updated 5 months ago
- Official repo for BLAP: Bootstrapping Language-Audio Pre-training for Music Captioning presented at ICASSP 2025☆16Nov 18, 2024Updated last year
- Hubert-based Forced Aligner☆56Mar 19, 2026Updated 4 months ago
- Training-Efficient Text-to-Music Generation with State-Space Modeling☆16Jan 31, 2026Updated 6 months ago
- This is the official code for ACM CIKM 2025 Paper: ParaStyleTTS: Toward Efficient and Robust Paralinguistic Style Control for Expressive …☆59Dec 21, 2025Updated 7 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The official repository TimeAudio, a comprehensive framework that incorporates fine-grained acoustic cues into LALMs with enhanced module…☆30Nov 18, 2025Updated 8 months ago
- A dataset of 173 progressive metal songs, in both GuitarPro and token formats, as per the specifications in DadaGP.☆18Nov 19, 2024Updated last year
- ☆22Jan 3, 2026Updated 7 months ago
- Phoneme Level Lyrics Alignment and Text-Informed Singing Voice Separation☆24Nov 8, 2021Updated 4 years ago
- ☆51Apr 30, 2026Updated 3 months ago
- ☆24Nov 16, 2025Updated 8 months ago
- M7-TTS: A Mini-Scale Multilingual and Multi-Dialect Text-to-Speech Language Model with Mimi codec and Multi Token Prediction☆20Mar 19, 2026Updated 4 months ago
- Realization for note segmentation by using hierarchical objective function☆14Jun 26, 2019Updated 7 years ago
- ☆28May 22, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official PyTorch implementation of CoverHunter☆44Nov 21, 2024Updated last year
- Power-Guided Grouped SRU for Real-Time Causal Audio-Visual Speech Separation☆26Jul 20, 2026Updated 2 weeks ago
- State-of-the-art continious audio tokenization☆41Mar 9, 2026Updated 4 months ago
- source code of EfficientTTS 2☆21Feb 18, 2024Updated 2 years ago
- Official source code of the INTERSPEECH 2023 paper: "Audio-Visual Speech Separation in Noisy Environments with a Lightweight Iterative Mo…☆20Sep 1, 2023Updated 2 years ago
- A repo that builds text to music datasets from scratch, used in MuseContorlLite [ICML2025]☆28May 20, 2025Updated last year
- Video Background Music Generation Using Unpaired Audio-Visual Data☆33Oct 8, 2024Updated last year
- [ICLR2026] FlexiCodec: A Dynamic Neural Audio Codec for Low Frame Rates☆51Jul 1, 2026Updated last month
- LEARNING A REPRESENTATION FOR COVER SONG IDENTIFICATION USING CONVOLUTIONAL NEURAL NETWORK. ICASSP2020☆54Jun 15, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for INTERSPEECH 2023 paper "mdctGAN: Taming transformer-based GAN for speech super-resolution with Modified DCT spectra"☆66Jun 3, 2023Updated 3 years ago
- Towards Fine-Grained Multi-Dimensional Speech Understanding: Data Pipeline, Benchmark, and Model☆25May 21, 2026Updated 2 months ago
- Generative Adaptive MIDI Extractor☆237Jun 14, 2026Updated last month
- ☆13Dec 18, 2017Updated 8 years ago
- Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation☆68Jun 16, 2026Updated last month
- Towards a general language-audio model for computational paralinguistic tasks☆31Dec 14, 2024Updated last year
- ☆34May 22, 2026Updated 2 months ago