Learnable STRF, from Riad et al. 2021 JASA
☆13Aug 21, 2021Updated 5 years ago
Alternatives and similar repositories for learnable-strf
Users that are interested in learnable-strf are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Automatically setup the AISHELL-4 and MSDWild dataset for usage with pyannote-database (and pyannote-audio)☆14Oct 22, 2025Updated 10 months ago
- Pytorch port of Google Research's LEAF Audio paper☆91May 19, 2021Updated 5 years ago
- ☆12Nov 7, 2024Updated last year
- 🎹 pyannote + 🗒 notebook = pyannotebook☆27Jun 12, 2023Updated 3 years ago
- Automatic speech annotator processing speech with voice activaty detection, overlapping speech detection, speaker diarization and automat…☆33Jun 14, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Companion repository for the paper "A Comparison of Metric Learning Loss Functions for End-to-End Speaker Verification" published at SLSP…☆61Oct 7, 2020Updated 5 years ago
- MicRank is a Learning to Rank neural channel selection framework where a DNN is trained to rank microphone channels.☆22Apr 8, 2021Updated 5 years ago
- Permutation invariant training in PyTorch☆13Oct 2, 2020Updated 5 years ago
- a wavelet transform for mapping behavior☆10May 4, 2024Updated 2 years ago
- ☆14Nov 28, 2022Updated 3 years ago
- Asteroid's filterbanks☆90Jan 12, 2025Updated last year
- A powerful tool to design any tensor factorization model and estimate the corresponding parameters☆12Sep 24, 2025Updated 11 months ago
- Official repo for the STRFNet system appeared in INTERSPEECH2020☆12Mar 6, 2021Updated 5 years ago
- ☆22Jun 30, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PyTorch implementation of "Nextformer: A ConvNeXt Augmented Conformer For End-To-End Speech Recognition"☆10Dec 15, 2022Updated 3 years ago
- ☆38Updated this week
- Clustering Analysis via Deep Generative Models With Mixture Models powered by@pytorch☆14Jan 6, 2022Updated 4 years ago
- A speech signal processing library in Python with emphasis on deep learning.☆31Apr 13, 2026Updated 4 months ago
- System that ranks 2nd in DCASE 2022 Challenge Task 5: Few-shot Bioacoustic Event Detection☆28Jul 6, 2022Updated 4 years ago
- Clean and modernized implementation of FastSpeech2/LightSpeech using IPA☆20Aug 16, 2024Updated 2 years ago
- ☆12Sep 12, 2024Updated last year
- A Python Library for Fundamental Frequency Estimation in Music Recordings☆55Jun 5, 2026Updated 2 months ago
- Unsupervised Voice Activity Detection by Modeling Source and System Information using Zero Frequency Filtering☆23Oct 19, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Simple sinc interpolation in PyTorch.☆15Jul 8, 2023Updated 3 years ago
- Audio activity detector based on per-channel energy normalization (PCEN)☆32Nov 16, 2018Updated 7 years ago
- Unofficial implementation of ConvNeXt-TTS powered by lightning☆18Oct 20, 2024Updated last year
- Text-to-dysarthric speech (TTDS) synthesis. An implementation using the Grad-TTS model with the TORGO database.☆14Mar 15, 2025Updated last year
- A pre-trained deep learning system for detecting bird flight calls in continuous recordings☆104Jun 17, 2024Updated 2 years ago
- ☆10Mar 22, 2023Updated 3 years ago
- Simple diarization model☆53Jun 13, 2025Updated last year
- melodic object transcription framework☆26Nov 15, 2017Updated 8 years ago
- Tensorflow 2.x implementation of the beta-TCVAE (arXiv:1802.04942).☆16Nov 12, 2019Updated 6 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Convolutional REpresenations for Music Analysis☆13Jul 5, 2016Updated 10 years ago
- Dissimilarity Matrix and Sounds from Timbre Space Representation of a Subtractive Synthesizer (Timbre, 2020)☆12Dec 17, 2021Updated 4 years ago
- ☆19Mar 22, 2024Updated 2 years ago
- Generate a false colour index spectrogram to visualise long-duration soundscape recordings☆32Mar 18, 2025Updated last year
- C++ version of pyannote audio overlapped speech detection pipeline☆13Feb 14, 2024Updated 2 years ago
- unofficial implementation of "CPTNN: CROSS-PARALLEL TRANSFORMER NEURAL NETWORK FOR TIME-DOMAIN SPEECH ENHANCEMENT"☆15Nov 14, 2023Updated 2 years ago
- PyTorch Implementation of SubSpectralNet - Using Sub-Spectrogram based Convolutional Neural Networks for Acoustic Scene Classification, a…☆21Feb 20, 2019Updated 7 years ago