gzhu06 / CacophonyLinks
Inference codebase for "Cacophony: An Improved Contrastive Audio-Text Model". Preprint: https://arxiv.org/abs/2402.06986
☆48Updated 2 weeks ago
Alternatives and similar repositories for Cacophony
Users that are interested in Cacophony are comparing it to the libraries listed below
Sorting:
- PyTorch implementation of the ICASSP-24 paper: "Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Superv…☆38Updated 2 years ago
- ☆45Updated last year
- Code for ICLR 2024 Paper: CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models