Implementing VGGVox for Speaker Identification on VoxCeleb1 dataset in PyTorch.
☆25Oct 15, 2020Updated 5 years ago
Alternatives and similar repositories for VGGVox-PyTorch
Users that are interested in VGGVox-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of the VGGVox network using TensorFlow.☆17Mar 20, 2026Updated 4 months ago
- acnn for text-independent speaker recognition☆10Feb 8, 2022Updated 4 years ago
- Speaker identification with VGGVox network☆84Nov 30, 2018Updated 7 years ago
- A curated list of awesome speaker recognition/verification papers, projects, datasets, and competition.☆15Aug 29, 2021Updated 4 years ago
- Text independent speaker recognition algorithm based on CNN☆24Aug 30, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- In defence of metric learning for speaker recognition☆1,171Apr 22, 2026Updated 3 months ago
- Predicting Political Instability and Social Conflicts Using Multimodal Data☆10Jun 6, 2016Updated 10 years ago
- A Pytorch implementation of triplet loss on VoxCeleb1☆12Oct 16, 2019Updated 6 years ago
- 分别在VCTK、AISHELL1 和 VoxCeleb1 三个标准公开数据集上对三种端到端声纹模型框架(Deep Speaker, RawNet, GE2E)进行实验比较。☆22Jun 23, 2020Updated 6 years ago
- SVHF-Net for Cross-modal binary matching☆32Aug 22, 2018Updated 7 years ago
- Pytorch implementation of "Towards Practical and Efficient Image-to-Speech Captioning with Vision-Language Pre-training and Multi-modal T…☆12Apr 29, 2026Updated 3 months ago
- Code for the paper:<LARNet:Lie Algebra Residual Network for Profile Face Recognition>(ICML2021)☆10Aug 19, 2021Updated 4 years ago
- DilatedSegNet: A Deep Dilated Segmentation Network for Polyp Segmentation☆13Oct 1, 2022Updated 3 years ago
- ☆11Sep 4, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Some codes and studies of regularization and normalization in GANs☆14Oct 14, 2021Updated 4 years ago
- Run commands on remote hosts, inspecting key indicators to manage infrastructure☆15Jan 29, 2026Updated 6 months ago
- The pytorch implementation for the paper of 'Cascaded Residual Attention Enhanced Road Extraction from Remote Sensing Images'☆14Dec 29, 2021Updated 4 years ago
- text-independent speaker identification☆12Apr 9, 2018Updated 8 years ago
- ☆16Mar 19, 2026Updated 4 months ago
- Code for paper NeuroGen: activation optimized image synthesis for discovery neuroscience.☆12Sep 24, 2023Updated 2 years ago
- Speaker Diarization using GRU in PyTorch☆11Aug 29, 2020Updated 5 years ago
- ☆12May 8, 2021Updated 5 years ago
- For IEEE ASRU(2025)☆15Jun 21, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆15Jul 15, 2019Updated 7 years ago
- ISIC-2018 Lesion Boundary Segmentation, Attribute Detection and Disease Classification☆14Jan 25, 2019Updated 7 years ago
- ☆19Jun 24, 2026Updated last month
- [ISBI'20] Complementary Network with Adaptive Receptive Fields for Melanoma Segmentation☆14Dec 21, 2020Updated 5 years ago
- ☆12Mar 23, 2021Updated 5 years ago
- ☆11May 4, 2020Updated 6 years ago
- The project tries to solve a speaker diarization problem using audio features, face recognition and video feature extraction from face im…☆16Feb 10, 2019Updated 7 years ago
- ☆26Jun 5, 2024Updated 2 years ago
- code for the paper "Cascade Attention Guided Residue GAN for Cross-Modal Translation"☆17Oct 31, 2020Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official repository for paper "Scene-agnostic Pose Regression for Visual Localization" (SPR), CVPR 2025☆34Mar 26, 2025Updated last year
- Demo for 2022 ICASSP☆64Jun 14, 2022Updated 4 years ago
- A repository comprising of code for generation of noisy speech data from clean data using deep learning methods☆16Jul 12, 2021Updated 5 years ago
- ☆20Dec 25, 2025Updated 7 months ago
- ☆19Jun 8, 2021Updated 5 years ago
- Course project for EE698R (2020-21 Sem 2). An X-Vector Based Speaker Diarization System with AutoEncoder based clustering method. Also su…☆16Jun 2, 2021Updated 5 years ago
- ☆18Apr 6, 2022Updated 4 years ago