[TOMM 2024] Automatic Lyric Transcription and Automatic Music Transcription from Multimodal Singing
☆28Aug 30, 2024Updated last year
Alternatives and similar repositories for SVT_SpeechBrain
Users that are interested in SVT_SpeechBrain are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code accompayning ISMIR23 paper; TriAD: Capturing harmonics with 3D convolutions☆20Jul 19, 2024Updated 2 years ago
- ☆14Feb 3, 2026Updated 5 months ago
- Frechet Audio Distance evaluation in PyTorch☆36Jun 9, 2023Updated 3 years ago
- Unconditional music synthesis using a diffusion model in the STFT domain☆12May 31, 2022Updated 4 years ago
- [MM 2022] MM-ALT: A Multimodal Automatic Lyric Transcription System (Oral, Top paper award)☆21Mar 16, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- High-Resolution Violin Transcription using Weak Labels☆41Oct 29, 2023Updated 2 years ago
- Extension of Sinsy-NG using deep learning models for voice conversion in order to synthesize good and realistic vocals.☆13Aug 14, 2020Updated 5 years ago
- This repo contains the source code of the first deep learning-base singing voice beat tracking system. It leverages WavLM and DistilHuBER…☆35Sep 4, 2022Updated 3 years ago
- Supplementary Materials of ISMIR 2022 paper "Analysis and detection of singing techniques in repertoires of J-POP solo singers" by Yuya Y…☆23Apr 23, 2024Updated 2 years ago
- Source code of paper "Adapting pretrained speech model for Mandarin lyrics transcription and alignment"☆19Dec 14, 2023Updated 2 years ago
- Template demonstrating how a manager may use Silver Bullet☆13Jul 7, 2023Updated 3 years ago
- ☆32Nov 25, 2023Updated 2 years ago
- ☆25Feb 20, 2024Updated 2 years ago
- ☆41Feb 1, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A dataset of 173 progressive metal songs, in both GuitarPro and token formats, as per the specifications in DadaGP.☆18Nov 19, 2024Updated last year
- A hackathon project to explore reworking the Mattermost Plugin API.☆11Aug 22, 2023Updated 2 years ago
- PyTorch implementation of DiffRoll, a diffusion-based generative automatic music transcription (AMT) model☆81Dec 6, 2023Updated 2 years ago
- [ISMIR 2022] Transfer Learning of wav2vec 2.0 for Automatic Lyric Transcription☆51May 7, 2024Updated 2 years ago
- ☆15Jul 16, 2026Updated last week
- This is the official implementation of our neural-network-based fast diffuse room impulse response generator (FAST-RIR) for generating r…☆12Nov 30, 2021Updated 4 years ago
- Tutorial covering Open Source tools for Source Separation.☆15Nov 12, 2021Updated 4 years ago
- Comparative Analysis of Graph Neural Networks for Node Regression task on Wiki-Squirrel dataset (Bachelor's Research Project)☆13Nov 6, 2025Updated 8 months ago
- [CVPR 2025] Official implementation of paper "Prosody-Enhanced Acoustic Pre-training and Acoustic-Disentangled Prosody Adapting for Movie…☆23Jun 6, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Generate a complete audio clip with music, intelligible speech, and sound effects from text in one pass.☆44May 27, 2026Updated last month
- A light digital audio workstation in JS☆14Jan 27, 2023Updated 3 years ago
- (Experimental) Predicting hand assignments in piano MIDI using neural networks☆13Oct 11, 2024Updated last year
- Code for "A diffusion-inspired training strategy for singing voice extraction in the waveform domain" (ISMIR 2022)☆17Feb 16, 2023Updated 3 years ago
- Node For Max Music experiments☆13Feb 15, 2018Updated 8 years ago
- The MIR-MLPop dataset and the official implementation of the paper "MIR-MLPop: A Multilingual Pop Music Dataset with Time-Aligned Lyrics …☆35Apr 22, 2024Updated 2 years ago
- ☆15Sep 24, 2022Updated 3 years ago
- A Large Dataset of Paired Guitar Audio Recordings and Tablatures☆25Sep 30, 2025Updated 9 months ago
- wake-up word emotion recognition [APSIPA 2022]☆17Nov 11, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- MAPS ( MIDI Aligned Piano Sounds ) dataset python api for machine learning☆11Jun 26, 2018Updated 8 years ago
- Implementation of Acoustic BPE (Shen et al., 2024), extended for RVQ-based Neural Audio Codecs☆76Dec 3, 2025Updated 7 months ago
- This repository provides the materials used in "Unsupervised Melody-to-Lyric Generation" by Yufei Tian, Anjali Narayan-Chen, Shereen Orab…☆11Jul 6, 2023Updated 3 years ago
- Deep learning for automatic mixing☆32Aug 29, 2024Updated last year
- 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production☆37Mar 10, 2022Updated 4 years ago
- ☆13Sep 23, 2021Updated 4 years ago
- Go-style channel and waitGroup for js to handle task queue.☆17Mar 26, 2025Updated last year