A light weight neural speaker embeddings extraction based on Kaldi and PyTorch.
☆136Jan 27, 2020Updated 6 years ago
Alternatives and similar repositories for pytorch-kaldi-neural-speaker-embeddings
Users that are interested in pytorch-kaldi-neural-speaker-embeddings are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- DropClass and DropAdapt - repository for the paper accepted to Speaker Odyssey 2020☆22Oct 29, 2020Updated 5 years ago
- VCTK multi-speaker tacotron for ICASSP 2020☆266Mar 29, 2022Updated 4 years ago
- Neural speaker recognition/verification system based on Kaldi and Tensorflow☆31Jun 30, 2020Updated 6 years ago
- Pytorch implementation of "Generalized End-to-End Loss for Speaker Verification"☆103Mar 18, 2019Updated 7 years ago
- Deep speaker embeddings in PyTorch, including x-vectors. Code used in this work: https://arxiv.org/abs/2007.16196☆322Nov 11, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implementation of Neural PLDA (NPLDA) model (A discriminative backend for Speaker Verification)☆99Apr 20, 2020Updated 6 years ago
- ☆35Apr 8, 2019Updated 7 years ago
- Tensorflow implementation of x-vector topology on top of Kaldi recipe☆118Nov 5, 2019Updated 6 years ago
- ☆46Oct 24, 2020Updated 5 years ago
- ☆37May 8, 2021Updated 5 years ago
- Speaker embedding(verification and recognition) using Pytorch☆369Jul 24, 2020Updated 6 years ago
- An Open Source Tools for Speaker Recognition☆636Aug 5, 2024Updated 2 years ago
- GPU accelerated implementation of i-vector extractor training using PyTorch. Requires Kaldi for feature extraction and UBM training. An e…☆63Oct 15, 2019Updated 6 years ago
- D3M - Dynamic Data Discrepancy Mitigation for Anti-spoofing - Implementation of work Dynamically Mitigating Data Discrepancy with Balance…☆30Feb 15, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Interface for Controllable Expressive Talking Machine☆40Sep 20, 2025Updated 11 months ago
- Deep Speaker: an End-to-End Neural Speaker Embedding System.☆941Apr 13, 2024Updated 2 years ago
- An implementation of "Investigation of enhanced Tacotron text-to-speech synthesis systems with self-attention for pitch accent language" …☆114Jun 19, 2020Updated 6 years ago
- Collection of self-supervised models for speaker and language recognition tasks.☆19Jan 18, 2022Updated 4 years ago
- Simple d-vector based Speaker Recognition (verification and identification) using Pytorch☆213Jul 17, 2020Updated 6 years ago
- Speaker embedding (d-vector) trained with GE2E loss☆290Jan 8, 2024Updated 2 years ago
- Keras implementation of SincNet (https://github.com/mravanelli/SincNet, https://arxiv.org/abs/1808.00158)☆12Aug 5, 2018Updated 8 years ago
- PyTorch implementation of a self-attentive speaker embedding☆17Sep 24, 2019Updated 6 years ago
- Time delay neural network (TDNN) implementation in Pytorch using unfold method☆207Nov 21, 2019Updated 6 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- In defence of metric learning for speaker recognition☆1,172Apr 22, 2026Updated 4 months ago
- PyTorch based speaker embedding model☆16Apr 13, 2024Updated 2 years ago
- PyTorch implementation of the Factorized TDNN (TDNN-F) from "Semi-Orthogonal Low-Rank Matrix Factorization for Deep Neural Networks" and …☆150Jan 6, 2020Updated 6 years ago
- ERISHA is a mulitilingual multispeaker expressive speech synthesis framework. It can transfer the expressivity to the speaker's voice for…☆44Dec 17, 2020Updated 5 years ago
- PyTorch implementation of "Generalized End-to-End Loss for Speaker Verification" by Wan, Li et al.☆600Jan 20, 2022Updated 4 years ago
- the Tensorflow version of multi-speaker TTS training with feedback constraint☆40Oct 12, 2020Updated 5 years ago
- [InterSpeech 2020] "AutoSpeech: Neural Architecture Search for Speaker Recognition" by Shaojin Ding*, Tianlong Chen*, Xinyu Gong, Weiwei …☆207Dec 8, 2022Updated 3 years ago
- Companion repository for the paper "A Comparison of Metric Learning Loss Functions for End-to-End Speaker Verification" published at SLSP…☆61Oct 7, 2020Updated 5 years ago
- Problem Agnostic Speech Encoder☆446Jul 6, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆51Feb 15, 2019Updated 7 years ago
- Code for DeCoAR (ICASSP 2020) and BERTphone (Odyssey 2020)☆104Nov 26, 2022Updated 3 years ago
- ☆24Jun 28, 2019Updated 7 years ago
- A WaveRNN implementation☆201Oct 14, 2019Updated 6 years ago
- SincNet is a neural architecture for efficiently processing raw audio samples.☆1,243Apr 28, 2021Updated 5 years ago
- A pure python module for reading and writing kaldi ark files☆267Mar 6, 2025Updated last year
- A pytorch implementation of xvector embedding☆79Mar 28, 2020Updated 6 years ago