Official repo for BLAP: Bootstrapping Language-Audio Pre-training for Music Captioning presented at ICASSP 2025
☆16Nov 18, 2024Updated last year
Alternatives and similar repositories for blap
Users that are interested in blap are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- "Enhancing Neural Audio Fingerprint Robustness to Audio Degradation for Music Identification" ISMIR2025☆38Sep 11, 2025Updated 10 months ago
- ☆22Jan 3, 2026Updated 7 months ago
- [ICASSP 2026] TinyMU: A Compact Audio Language Model for Music Understanding☆38Apr 20, 2026Updated 3 months ago
- ☆33Dec 23, 2025Updated 7 months ago
- MiRA (Music Replication Assessment) tool is a model-independent open evaluation method based on four diverse audio music similarity metri…☆35Nov 14, 2025Updated 8 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆19Sep 20, 2025Updated 10 months ago
- Codes for ICASSP 2024 paper: BEAST: Online Joint Beat and Downbeat Tracking Based on Streaming Transformer. An online beat tracking syste…☆44Sep 11, 2024Updated last year
- Project page of "2026-ICLR Echo: Towards Advanced Audio Comprehension via Audio-Interleaved Reasoning"☆16Mar 26, 2026Updated 4 months ago
- MuChoMusic is a benchmark for evaluating music understanding in multimodal audio-language models.☆46Dec 3, 2024Updated last year
- MelodySim: Measuring Melody-aware Music Similarity for Plagiarism Detection☆29May 29, 2025Updated last year
- LLM-Codec: Neural Audio Codec Meets Language Model Objectives☆23May 3, 2026Updated 3 months ago
- ☆23Jun 18, 2026Updated last month
- ☆10Sep 25, 2024Updated last year
- Official repository for the paper - SLAP: Siamese Language-Audio Pretraining without negative samples for Music Understanding☆63Sep 25, 2025Updated 10 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆15Feb 6, 2026Updated 6 months ago
- [ICML2026] AudioMosaic: Contrastive Masked Audio Representation Learning☆23May 15, 2026Updated 2 months ago
- Audio Entailment: Deductive Reasoning for Audio Understanding☆17Dec 10, 2024Updated last year
- [ACM MM 2025] AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation☆24Oct 28, 2025Updated 9 months ago
- Code for ICASSP 2024 Paper: RECAP: Retrieval-Augmented Audio Captioning☆16Jun 23, 2024Updated 2 years ago
- Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation☆68Jun 16, 2026Updated last month
- This is the official implementation of RL-Chord (TNNLS).☆13Jan 2, 2024Updated 2 years ago
- Code for ICLR 2024 Paper: CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models☆23Jul 10, 2024Updated 2 years ago
- An All-in-One Speech, Sound, Music Codec with Single Nested Codebook☆28Oct 11, 2025Updated 9 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICASSP2025] Official code for VoiceDiT: Dual-Condition Diffusion Transformer for Environment-Aware Speech Synthesis☆52Apr 9, 2025Updated last year
- ☆29Apr 6, 2026Updated 4 months ago
- A toolbox for objective evaluation in symbolic music generation.☆100Aug 2, 2021Updated 5 years ago
- Official source codes of coco-mulla☆36Mar 21, 2024Updated 2 years ago
- This is the official implementation of EmoMusicTV (TMM).☆26Jan 15, 2024Updated 2 years ago
- Frechet Audio Distance evaluation in PyTorch☆36Jun 9, 2023Updated 3 years ago
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated 11 months ago
- Explaining audio differences using language☆16Feb 11, 2025Updated last year
- Official implementation of TISDiSS, a scalable framework for discriminative source separation.☆16Jul 31, 2026Updated last week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Zero-Shot Blind Audio Bandwidth Extension☆27May 25, 2023Updated 3 years ago
- Official Repository for "Music Source Restoration"☆31Jun 1, 2025Updated last year
- ☆30Apr 22, 2024Updated 2 years ago
- ☆131Jul 23, 2026Updated 2 weeks ago
- Musical Word Embedding for Music Tagging and Retrieval [IEEE TASLP]☆29Apr 23, 2024Updated 2 years ago
- A curated list of models, benchmarks, tools and guides for audio editing☆35Updated this week
- ☆15Jan 9, 2026Updated 7 months ago