Official Pytorch implementation of the "A Model You Can Hear: Audio Identification with Playable Prototypes" paper
☆37Aug 8, 2022Updated 4 years ago
Alternatives and similar repositories for a-model-you-can-hear
Users that are interested in a-model-you-can-hear are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Supervised and unsupervised Concept-based explanation of pretrained music classifiers☆12Jul 27, 2023Updated 3 years ago
- DenseNets for the detection of singing birds in audio files☆19Nov 15, 2017Updated 8 years ago
- CNN-based singing voice detection experiments☆37Apr 25, 2018Updated 8 years ago
- A versatile, easily configurable vocoder software in MATLAB, for research purposes☆15Apr 9, 2021Updated 5 years ago
- Supplementary code for the experiments described in the 2021 ISMIR submission: Leveraging Hierarchical Structures for Few Shot Musical In…☆41Aug 12, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Website for the ISMIR 2023 Tutorial: Few-shot and Zero-shot Learning for MIR☆30Jan 3, 2023Updated 3 years ago
- ☆37May 22, 2026Updated 4 months ago
- ☆12May 12, 2016Updated 10 years ago
- An implementation of the Contrast Predictive Coding (CPC) method to train audio features in an unsupervised fashion.☆10Feb 22, 2022Updated 4 years ago
- My implementation of Epoch-Synchronous Overlap-Add method for time stretching and pitch shifting.☆10Jan 25, 2020Updated 6 years ago
- Voice100 includes neural TTS/ASR models. Inference of Voice100 is low cost as its models are tiny and only depend on CNN without autoregr…☆28Nov 23, 2023Updated 2 years ago
- Code and data repository for ISMIR 2019 paper: MIDI–SHEET MUSIC ALIGNMENT USING BOOTLEG SCORE SYNTHESIS☆12Mar 1, 2022Updated 4 years ago
- Generation tool for offset-resistant audio adversarial examples against Deepspeech☆10Oct 5, 2020Updated 5 years ago
- Bias Tests for Voice Technologies (bt4vt)☆11Jun 16, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- source code of "End-to-end Music Remastering System Using Self-supervised and Adversarial Training"☆47Sep 7, 2023Updated 3 years ago
- Experimenting with Lapped Transforms Jupyter Notebook☆14Jun 13, 2025Updated last year
- [ISMIR 2023] LyricWhiz: Robust Multilingual Zero-shot Lyrics Transcription by Whispering to ChatGPT☆56Nov 20, 2023Updated 2 years ago
- End-to-end waveform utterance enhancement for direct evaluation metrics optimization by fully convolutional neural networks (TASLP 2018)☆18Jul 12, 2019Updated 7 years ago
- ☆23Aug 30, 2022Updated 4 years ago
- ICLR 2019 Paper, "Characterizing Audio Adversarial Examples using Temporal Dependency".☆11Apr 3, 2019Updated 7 years ago
- ☆21Dec 17, 2018Updated 7 years ago
- Streaming source separation for music and speech files, using the Open-Unmix LSTM architecture.☆21Dec 8, 2022Updated 3 years ago
- Source of my website: tech blog, tech talks, fpv handbook and fpv builds☆13Sep 21, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- This repository contains the code for the paper "Exploiting Foundation Models and Speech Enhancement for Parkinson's Disease Detection fr…☆12Dec 19, 2025Updated 9 months ago
- official implementation of paper ExPO: Explainable Phonetic Trait-Oriented Network for Speaker Verification☆15Mar 14, 2025Updated last year
- melodic object transcription framework☆26Nov 15, 2017Updated 8 years ago
- ☆11Apr 8, 2016Updated 10 years ago
- Crawled from FreeMidi.org, MIDI files library including over twenty thousand files!☆33Jun 6, 2020Updated 6 years ago
- Python phase-vocoder implementation with pitch shifting and formant correction☆15Feb 17, 2022Updated 4 years ago
- TAPE: An End-to-End Timbre-Aware Pitch Estimator☆24Nov 25, 2023Updated 2 years ago
- Repository for DNN training, theory to practice, part of the Large Scale Machine Learning class at Mines Paritech☆11Mar 11, 2022Updated 4 years ago
- ☆88Jan 29, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Visual musicology - koncon Master project☆16Dec 7, 2023Updated 2 years ago
- Algorithms to automatically recognize guitar effects and retrieve their parameters for timbre reproduction☆27Aug 31, 2022Updated 4 years ago
- ☆27Mar 5, 2018Updated 8 years ago
- Baseline scripts for AVEC 2019, Depression Detection Sub-challenge☆16Jul 11, 2019Updated 7 years ago
- ☆16Jun 17, 2021Updated 5 years ago
- This is the codebase for defense framework described in USENIX '21 paper "WaveGuard: Understanding and Mitigating Audio Adversarial Examp…☆20Oct 20, 2021Updated 4 years ago
- Depression Recognition☆11Mar 11, 2024Updated 2 years ago