Speaker change detection using SincNet and an LSTM/Transformer
☆57May 26, 2025Updated last year
Alternatives and similar repositories for speaker-change-detection
Users that are interested in speaker-change-detection are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Online streaming speaker change detection model in Pytorch☆44Apr 14, 2023Updated 3 years ago
- ☆15Jul 11, 2022Updated 4 years ago
- Both audio-only and audio-visual speaker diarization datasets are listed here.☆16Feb 22, 2023Updated 3 years ago
- Paper: https://arxiv.org/abs/1702.02285☆64Dec 19, 2018Updated 7 years ago
- Automatically setup the AISHELL-4 and MSDWild dataset for usage with pyannote-database (and pyannote-audio)☆15Oct 22, 2025Updated 9 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Artie Bias Corpus: an audio corpus + code for detecting demographic bias☆20Jul 21, 2020Updated 6 years ago
- ☆11May 4, 2020Updated 6 years ago
- ☆327Jun 14, 2024Updated 2 years ago
- Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment☆14Feb 5, 2025Updated last year
- The VoxTube dataset official repository☆71Feb 14, 2024Updated 2 years ago
- CLASP: Contrastive Language-Speech Pretraining for Multilingual Multimodal Information Retrieval☆13Jun 27, 2025Updated last year
- ☆37Jan 6, 2026Updated 6 months ago
- sherpa with mlx☆15Aug 2, 2025Updated 11 months ago
- [ICASSP 2025] AnCoGen: Analysis, Control and Generation of Speech with a Masked Autoencoder☆14Mar 11, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An N-gram punctuator for Chinese and English.☆18Oct 14, 2025Updated 9 months ago
- Predicts the level of noise and reverberation on your audiofiles☆191May 23, 2026Updated 2 months ago
- ☆12Jun 14, 2024Updated 2 years ago
- This repo is for the SPL paper "Auto-Tuning Spectral Clustering for Speaker Diarization Using Normalized Maximum Eigengap"☆125Apr 8, 2022Updated 4 years ago
- ☆12Jun 14, 2022Updated 4 years ago
- The official Pytorch implementation of "Frame-wise streaming end-to-end speaker diarization with non-autoregressive self-attention-based …☆183May 7, 2026Updated 2 months ago
- Thai Grapheme to Phoneme (G2P) Wiktionary Corpus☆13Jul 25, 2022Updated 4 years ago
- This is the code and dataset repo for Interspeech 2024 paper "Target conversation extraction: Source separation using turn-taking dynamic…☆58Aug 15, 2025Updated 11 months ago
- ☆22Jan 3, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Clustering-based methods for overlapping diarization☆84Jan 12, 2024Updated 2 years ago
- This repository contains the code for the paper "voc2vec: A Foundation Model for Non-Verbal Vocalization", accepted at ICASSP 2025.☆58Apr 14, 2025Updated last year
- An tensorflow implementation of ghostvlad for speaker recognition☆15May 2, 2019Updated 7 years ago
- Official implement of "Dual-stream Time-Delay Neural Network with Dynamic Global Filter for Speaker Verification" in PyTorch☆41Aug 31, 2023Updated 2 years ago
- ☆27Sep 10, 2025Updated 10 months ago
- A lightweight library to compute Diarization Error Rate (DER).☆62Jan 14, 2026Updated 6 months ago
- This repository contains source codes for SoftCTC. Original paper can be found here: https://arxiv.org/abs/2212.02135☆19Mar 7, 2023Updated 3 years ago
- Torch implementation of NANSY, Neural Analysis and Synthesis, arXiv:2110.14513☆64Feb 13, 2023Updated 3 years ago
- Simplified diarization pipeline using some pretrained models - audio file to diarized segments in a few lines of code☆158May 2, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- LIGHTVOC AN UPSAMPLING-FREE GAN VOCODER BASED ON CONFORMER AND INVERSE SHORT-TIME FOURIER TRANSFORM☆18May 17, 2024Updated 2 years ago
- An ODE-based generative neural vocoder using Rectified Flow☆58Apr 29, 2023Updated 3 years ago
- code for Towards Data Science article on prompt-loss-weight☆11Jun 4, 2025Updated last year
- ☆15Apr 16, 2026Updated 3 months ago
- This repository contains prompts & best practices to annotate audio clips with a very high degree of details using Audio-Language-Models☆35Oct 13, 2024Updated last year
- ☆17Apr 12, 2021Updated 5 years ago
- Python wrapper for OpenFST and its extensions from Kaldi. Also support reading/writing ark/scp files☆56Apr 9, 2026Updated 3 months ago