Source code of paper "Adapting pretrained speech model for Mandarin lyrics transcription and alignment"
☆19Dec 14, 2023Updated 2 years ago
Alternatives and similar repositories for LyricAlignment
Users that are interested in LyricAlignment are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repo contains the source code of the first deep learning-base singing voice beat tracking system. It leverages WavLM and DistilHuBER…☆35Sep 4, 2022Updated 3 years ago
- [ISMIR 2023] LyricWhiz: Robust Multilingual Zero-shot Lyrics Transcription by Whispering to ChatGPT☆56Nov 20, 2023Updated 2 years ago
- [ICASSP 2025] AnCoGen: Analysis, Control and Generation of Speech with a Masked Autoencoder☆15Mar 11, 2025Updated last year
- singing voice with annotations of vocal onsets, based on the matched MIDI from http://colinraffel.com/projects/lmd/☆20Dec 30, 2019Updated 6 years ago
- ☆12Nov 18, 2020Updated 5 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- PyTorch implementation of Sequence Transduction with Recurrent Neural Networks (RNN-T) speech recognition paper☆16Mar 4, 2022Updated 4 years ago
- This is a subset of the DALI set consisting of 240 polyphonic recordings that is used to benchmark lyrics transcription evaluation.☆12Nov 30, 2021Updated 4 years ago
- [CVPR'25] Attention IoU: Examining Biases in CelebA using Attention Maps☆13Mar 26, 2025Updated last year
- [TOMM 2024] Automatic Lyric Transcription and Automatic Music Transcription from Multimodal Singing☆28Aug 30, 2024Updated 2 years ago
- Code for the paper "Songs Across Borders: Singable and Controllable Neural Lyric Translation"☆26Feb 3, 2026Updated 7 months ago
- ☆12Nov 7, 2024Updated last year
- ☆25Jun 13, 2024Updated 2 years ago
- Repo of the paper "Towards Building an End-to-End Multilingual Automatic Lyrics Transcription Model""☆15Jun 28, 2024Updated 2 years ago
- wake-up word emotion recognition [APSIPA 2022]☆17Nov 11, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆19Aug 23, 2026Updated last week
- [ASRU 2023] Code of paper SALT: Distinguishable Speaker Anonymization Through Latent Space Transformation☆23Aug 13, 2024Updated 2 years ago
- A Roman Numeral Analysis Network with Synthetic Training Examples and Additional Tonal Tasks☆51Feb 11, 2024Updated 2 years ago
- version 4.x of the Princeton Geniza Project☆13Updated this week
- ☆24Nov 16, 2025Updated 9 months ago
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- Code accompayning ISMIR23 paper; TriAD: Capturing harmonics with 3D convolutions☆20Jul 19, 2024Updated 2 years ago
- GenerationMania: Generate IIDX-style rhythm action game stages☆12Aug 3, 2019Updated 7 years ago
- ☆14Jan 28, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Supplementary Materials of ISMIR 2022 paper "Analysis and detection of singing techniques in repertoires of J-POP solo singers" by Yuya Y…☆23Apr 23, 2024Updated 2 years ago
- The MIR-MLPop dataset and the official implementation of the paper "MIR-MLPop: A Multilingual Pop Music Dataset with Time-Aligned Lyrics …☆35Apr 22, 2024Updated 2 years ago
- ☆12Feb 20, 2025Updated last year
- A unified model for zero-shot singing voice conversion and synthesis☆22Nov 30, 2022Updated 3 years ago
- Official repository for the paper "AudioMAE++: learning better masked audio representations with SwiGLU FFNs"☆15Apr 30, 2026Updated 4 months ago
- Official Repository of IJCAI 2024 Paper: "BATON: Aligning Text-to-Audio Model with Human Preference Feedback"☆32Mar 4, 2025Updated last year
- ☆19Mar 27, 2023Updated 3 years ago
- Robust Singing Voice Transcription and MIDI Extraction☆126Nov 20, 2024Updated last year
- ☆18May 4, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implementation of SawSing (ISMIR'22)☆275Aug 28, 2022Updated 4 years ago
- Synthesized singing voice demos of WeSinger 2 paper.☆26Feb 20, 2023Updated 3 years ago
- 東北イタコ歌唱データベースの最新ラベルデータ☆24Jul 1, 2021Updated 5 years ago
- ☆13Nov 2, 2020Updated 5 years ago
- Phoneme Level Lyrics Alignment and Text-Informed Singing Voice Separation☆24Nov 8, 2021Updated 4 years ago
- SOFA: Singing-Oriented Forced Aligner☆237Updated this week
- Pytorch implementation of BigVSAN☆203Dec 9, 2025Updated 8 months ago