An attempt to replicate the results of [1706.08612] VoxCeleb: a large-scale speaker identification dataset
☆12Dec 11, 2019Updated 6 years ago
Alternatives and similar repositories for VoxCeleb
Users that are interested in VoxCeleb are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Speaker identification with VGGVox network☆84Nov 30, 2018Updated 7 years ago
- The project is related to the development of labs for the ITMO Speaker Recognition Course.☆16Jul 3, 2026Updated last month
- Score calibration for speaker verification☆25Dec 13, 2019Updated 6 years ago
- VGGVox models for Speaker Identification and Verification trained on the VoxCeleb (1 & 2) datasets☆402Feb 4, 2019Updated 7 years ago
- VoxCeleb plugin for pyannote.database☆30Aug 4, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Multiobjective Optimization Training of PLDA for Speaker Verification☆10Jun 14, 2018Updated 8 years ago
- Latest PyTorch Implementation of DeltaGRU & DeltaLSTM that Exploits Temporal Sparsity in Sequential Data☆18Sep 30, 2023Updated 2 years ago
- It uses GMM to train a gender detector model. The testing has been done on subset of Google's AudioSet corpus.☆19Jun 14, 2017Updated 9 years ago
- Gated CNN☆10Jul 17, 2019Updated 7 years ago
- pytorch maml with Multi-GPUs, fast and simplest implementation☆13Dec 4, 2020Updated 5 years ago
- ☆12Oct 17, 2024Updated last year
- kaldi based x-vector trained on Cn-Celeb☆13Sep 22, 2020Updated 5 years ago
- Keras framework for speech enhancement using relativistic GANs☆52Jun 24, 2020Updated 6 years ago
- neural network and loss for asv implemented by PyTorch. (Triplet loss, LMCL, Angular Loss, Softmax)☆21Oct 23, 2019Updated 6 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- A Pytorch implementation of triplet loss on VoxCeleb1☆12Oct 16, 2019Updated 6 years ago
- Text independent speaker recognition algorithm based on CNN☆24Aug 30, 2025Updated 11 months ago
- The Pytorch implementation of Network-In-Network☆11Jun 7, 2018Updated 8 years ago
- The code used to create the ARCA23K and ARCA23K-FSD datasets☆16Nov 9, 2021Updated 4 years ago
- The 1st place solution for AutoSpeech 2019.☆17Jun 9, 2020Updated 6 years ago
- Text-to-Speech Synthesis by Generating Spectrograms using Generative Adversarial Network☆10Dec 12, 2018Updated 7 years ago
- ☆12May 22, 2022Updated 4 years ago
- A recipe for creating a Speaker Identification system built on Kaldi.☆15Jan 2, 2020Updated 6 years ago
- ☆13Jan 10, 2017Updated 9 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Wav2kws is keyword spotting (KWS) based on Wav2Vec 2.0. This model shows state-of-the-art in Google Speech Commands datasets V1 and V2.☆13Jun 11, 2021Updated 5 years ago
- Deep Learning Projects with PyTorch [video], published by Packt☆12Jan 15, 2021Updated 5 years ago
- A collection of basic text processing modules focused on Gujarati☆10Oct 24, 2017Updated 8 years ago
- Adversarial examples on keras and tensorflow☆12Apr 5, 2017Updated 9 years ago
- An extension of thu-spmi/CAT which contains a full-fledged implementation of CTC-CRF for Tensorflow.☆12Jul 5, 2021Updated 5 years ago
- Tensorflow implementation of "Generalized End-to-End Loss for Speaker Verification"☆365Oct 9, 2021Updated 4 years ago
- ☆12Mar 9, 2023Updated 3 years ago
- Speaker embedding(verification and recognition) using Tensorflow with Kaldi☆41Sep 18, 2017Updated 8 years ago
- ☆15Jan 24, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code and dataset release for "PACS: A Dataset for Physical Audiovisual CommonSense Reasoning" (ECCV 2022)☆18Dec 20, 2022Updated 3 years ago
- Utterance-level Aggregation For Speaker Recognition In The Wild☆372Mar 24, 2023Updated 3 years ago
- This repository contains the video files (download links) and corresponding annotations used in the paper "Long-Term Face Tracking for Cr…☆14Dec 18, 2020Updated 5 years ago
- Neural speaker recognition/verification system based on Kaldi and Tensorflow☆31Jun 30, 2020Updated 6 years ago
- A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统☆14Mar 18, 2019Updated 7 years ago
- Estimating the Age, Height, and Gender of a speaker with their speech signal.☆15Sep 19, 2022Updated 3 years ago
- 以音素建模构建NN-CTC声学模型☆16May 14, 2019Updated 7 years ago