Using speaker embedding for diarization in PyTorch
☆17Aug 29, 2020Updated 6 years ago
Alternatives and similar repositories for pytorch_speaker_embedding_for_diarization
Users that are interested in pytorch_speaker_embedding_for_diarization are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- For our Smart Media Player (detecting time period(s) inside audio/video during which specific person(s) is/are speaking) project☆18Feb 25, 2020Updated 6 years ago
- kaldi based x-vector trained on Cn-Celeb☆13Sep 22, 2020Updated 6 years ago
- ☆11Sep 4, 2023Updated 3 years ago
- Speaker Diarization using GRU in PyTorch☆11Aug 29, 2020Updated 6 years ago
- Neural network based similarity scoring for diarization (pytorch implementation of "LSTM based Similarity Measurement with Spectral Clust…☆43Oct 21, 2020Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Speaker diarization and speech to text☆14Dec 17, 2020Updated 5 years ago
- The implementation of "End-to-End Neural Speaker Diarization with an Iterative Adaptive Attractor Estimation", which is accepted by Neura…☆11Aug 27, 2023Updated 3 years ago
- Kaldi style neural network training in pytorch for use in place of nnet3 in Kaldi.☆26Jul 25, 2024Updated 2 years ago
- Official PyTorch implementation of "t-EER: Parameter-Free Tandem Evaluation Metric of Countermeasures and Biometric Comparators"☆14Sep 25, 2023Updated 2 years ago
- Tensorflow Implementation for "Pre-trained Deep Convolution Neural Network Model With Attention for Speech Emotion Recognition"☆10Dec 19, 2021Updated 4 years ago
- Example python scripts to evaluate various ASR methods☆11Dec 22, 2021Updated 4 years ago
- This repository created for the NHN ASR hackathon competition.☆12Sep 20, 2023Updated 3 years ago
- The project tries to solve a speaker diarization problem using audio features, face recognition and video feature extraction from face im…☆16Feb 10, 2019Updated 7 years ago
- Tutorial session material of Pytest in PyCon KR 2019☆10Jul 22, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 책 읽어주는 딥러닝을 보고 나도 만들고 싶어져서 공부하며 만드는 repository입니다.☆10Dec 8, 2022Updated 3 years ago
- Create Interactive Dashboards With Streamlit in Python☆19Aug 12, 2020Updated 6 years ago
- ☆12Jun 14, 2022Updated 4 years ago
- ☆17May 23, 2025Updated last year
- [ASRU 2023] Code of paper SALT: Distinguishable Speaker Anonymization Through Latent Space Transformation☆23Aug 13, 2024Updated 2 years ago
- ☆11Mar 12, 2019Updated 7 years ago
- CDER (Conversational Diarization Error Rate) Scoring Tool☆22Sep 13, 2022Updated 4 years ago
- Materials for the Hugging Face Diffusion Models Course☆14Sep 22, 2025Updated last year
- ☆15Jul 15, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Applying Deep Reinforcement Learning for dialogue generation. aka chatbot☆13Apr 30, 2017Updated 9 years ago
- 用于存放自然语言处理相关的代码。Store code related to NLP (Natural Language Processing).☆13Sep 19, 2019Updated 7 years ago
- Langchain_CrewAI_Gemini - An Gemini AI powered AI Agent (Multi-Agent) Project.☆14Mar 24, 2024Updated 2 years ago
- Tools for downloading VoxCeleb2 dataset☆35Mar 16, 2024Updated 2 years ago
- ☆18Jun 8, 2019Updated 7 years ago
- Multi-Stage Face-Voice Association Learning with Keynote Speaker Diarization (ACM MM 2024)☆22Jul 25, 2024Updated 2 years ago
- FEERCI: A Package for Fast non-parametric confidence intervals for Equal Error Rates☆12Mar 13, 2024Updated 2 years ago
- Advances in audio anti-spoofing and deepfake detection using graph neural networks and self-supervised learning☆23Aug 20, 2023Updated 3 years ago
- Speaker Identification using Neural Net.☆20Jul 30, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- PyTorch based speaker embedding model☆16Apr 13, 2024Updated 2 years ago
- A Flutter plugin for inference of Pytorch models (https://pub.dev/packages/torch_mobile). Supports image classification on Android.☆26Oct 17, 2019Updated 6 years ago
- Large Language-and-Vision Assistant for BioMedicine, built towards multimodal GPT-4 level capabilities.☆10Nov 29, 2023Updated 2 years ago
- PHO-LID: A Unified Model to Incorporate Acoustic-Phonetic and Phonotactic Information for Language Identification☆21Aug 24, 2023Updated 3 years ago
- Análisis de Datos (pregrado)☆12Jan 8, 2021Updated 5 years ago
- An end-to-end MATLAB toolkit for completely unsupervised Speaker Diarization using state-of-the-art algorithms.☆15Dec 22, 2015Updated 10 years ago
- Hospital simulator with pedestrians and robot☆15Oct 20, 2024Updated last year