Cross attentive pooling for speaker verification (IEEE SLT, 2021)
☆12Dec 14, 2020Updated 5 years ago
Alternatives and similar repositories for CAP
Users that are interested in CAP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Baseline for the Spoofing-aware Speaker Verification Challenge 2022☆68May 3, 2022Updated 4 years ago
- Pytorch implementation of Meta-Learning for Short Utterance Speaker Recognition with Imbalance Length Pairs (Interspeech, 2020)☆73Sep 16, 2020Updated 5 years ago
- Pytorch implementation of RawNeXt: Speaker verification system for variable-duration utterance with deep layer aggregation and dynamic sc…☆25Jun 22, 2022Updated 4 years ago
- ☆21Apr 6, 2021Updated 5 years ago
- Score calibration for speaker verification☆25Dec 13, 2019Updated 6 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- This is my speaker recognition implementation based on the x-vector system described in "X-Vectors: Robust DNN Embeddings for Speaker Rec…☆11Jul 23, 2026Updated last month
- Python3 code for the IEEE SPL paper "Auto-Tuning Spectral Clustering for SpeakerDiarization Using Normalized Maximum Eigengap"☆11Apr 6, 2020Updated 6 years ago
- Collection of self-supervised models for speaker and language recognition tasks.☆19Jan 18, 2022Updated 4 years ago
- Multi-speaker & Multi-style TTS☆27Jul 3, 2024Updated 2 years ago
- Augmentation adversarial training for self-supervised speaker recognition☆77Aug 15, 2021Updated 5 years ago
- Official PyTorch implementation of "t-EER: Parameter-Free Tandem Evaluation Metric of Countermeasures and Biometric Comparators"☆14Sep 25, 2023Updated 2 years ago
- UPC Deep Learning for Speech and Language 2018☆17Feb 26, 2018Updated 8 years ago
- PyTorch Implementation for the paper "DisCont: Self-Supervised Visual Attribute Disentanglement using Context Vectors" (ECCVW'20).☆15Oct 23, 2020Updated 5 years ago
- ML/DL training workshops for EEE undergrads☆13Jan 16, 2019Updated 7 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- An UWP client software for ASRT speech recognition system. 一个可用于ASRT语音识别系统的UWP客户端软件☆12Oct 23, 2019Updated 6 years ago
- Official release of pretrained models and codes for 'Golden Gemini Is All You Need: Finding the Sweet Spots for Speaker Verification'☆21Jan 20, 2025Updated last year
- ChatTube: A Retrieval QA System to Youtube Videos☆10Jun 6, 2023Updated 3 years ago
- This is an official PyTorch code for our accepted paper "When All We Need is a Piece of the Pie: A Generic Framework for Optimizing Two-w…☆15Jul 7, 2022Updated 4 years ago
- Real Time Chat Application☆14Dec 20, 2022Updated 3 years ago
- Official repository for RawNet, RawNet2, and RawNet3☆407Mar 21, 2024Updated 2 years ago
- In defence of metric learning for speaker recognition☆1,176Apr 22, 2026Updated 4 months ago
- Development Toolkit for the VoxCeleb Speaker Recognition Challenge 2021☆19Jul 21, 2021Updated 5 years ago
- Gated graph convolutional recurrent neural networks code used in:☆17Jul 10, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code and dataset for Polyglot Prompting: Multilingual Multitask Prompt Training.☆18Dec 7, 2022Updated 3 years ago
- ☆35Nov 24, 2024Updated last year
- CLI tool to record how much time it takes to import each dependency in a Python project☆11Mar 24, 2022Updated 4 years ago
- Codes for paper -- Towards Training Explainable Singing Quality Assessment Network with Augmented Data☆16Dec 7, 2021Updated 4 years ago
- It is fine-tune the GPT-Neo model for Thai language.☆12Jun 30, 2021Updated 5 years ago
- ☆10Apr 10, 2019Updated 7 years ago
- Metric Learning (npair loss & angular loss) on mnist and Visualizing by t_SNE☆35Feb 15, 2023Updated 3 years ago
- commandline download of ARD videos☆16Jan 26, 2021Updated 5 years ago
- SASV2 baseline, a track on ASVspoof5 phase2 challenge☆28Nov 12, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Cosine Similary Search in ElasticSearch + FAISS GPU☆12Mar 24, 2022Updated 4 years ago
- Implementation and Deployment of Multilingual Custom Keyword Spotting Running in Real-time on an Edge Device.☆11Apr 27, 2023Updated 3 years ago
- MTGAN: Speaker Verification through Multitasking Triplet Generative Adversarial Networks☆19Feb 29, 2020Updated 6 years ago
- [IJCAI2022] Unsupervised Voice-Face Representation Learning by Cross-Modal Prototype Contrast☆22Oct 25, 2023Updated 2 years ago
- Python toolkit for speech processing☆72Aug 22, 2026Updated last week
- Speaker Verification using Pytorch☆13May 23, 2024Updated 2 years ago
- A Model (maybe an app) that translates the audio of a video from one language to another language, cloning the voice of original video wi…☆17May 19, 2025Updated last year