Speaker Identification using Neural Net.
☆20Jul 30, 2024Updated 2 years ago
Alternatives and similar repositories for speaker-identification
Users that are interested in speaker-identification are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆12Jun 14, 2022Updated 4 years ago
- Create speaker voiceprints from a few seconds of audio. And, identify individuals in real-time streaming or recorded conversations.☆17Feb 4, 2019Updated 7 years ago
- Official PyTorch implementation of "t-EER: Parameter-Free Tandem Evaluation Metric of Countermeasures and Biometric Comparators"☆14Sep 25, 2023Updated 2 years ago
- ☆12Aug 23, 2019Updated 6 years ago
- Using speaker embedding for diarization in PyTorch☆17Aug 29, 2020Updated 5 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Region-Based Optimization in Continual Learning for Audio Deepfake Detection☆14Dec 17, 2024Updated last year
- acin_academy☆15Jun 22, 2026Updated last month
- Official implementation of SBNet as described in "Single-branch Network for Multimodal Training".☆13Aug 28, 2023Updated 2 years ago
- ☆14Jan 7, 2023Updated 3 years ago
- [ASRU 2023] Code of paper SALT: Distinguishable Speaker Anonymization Through Latent Space Transformation☆23Aug 13, 2024Updated 2 years ago
- Bimodal Adaptive Feature Fusion Network for Person Verification☆20Jul 30, 2022Updated 4 years ago
- self ensemble label correction☆17Jul 29, 2022Updated 4 years ago
- ☆14Dec 21, 2024Updated last year
- Leveraging BERT to Improve Spoken Language Identification☆17Nov 22, 2022Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆11Nov 5, 2021Updated 4 years ago
- Evaluation script for VoxMovies dataset in PyTorch☆23Jan 12, 2024Updated 2 years ago
- Advances in audio anti-spoofing and deepfake detection using graph neural networks and self-supervised learning☆23Aug 20, 2023Updated 2 years ago
- Voice Face Association Learning Paper List☆17May 20, 2023Updated 3 years ago
- PHO-LID: A Unified Model to Incorporate Acoustic-Phonetic and Phonotactic Information for Language Identification☆21Aug 24, 2023Updated 2 years ago
- ☆13May 1, 2026Updated 3 months ago
- This is the official train-dev-test release of the Interspeech2024 Discrete Speech Representation Challenge.☆32Jan 26, 2024Updated 2 years ago
- Discriminative Training of VBx Diarization☆28Sep 23, 2024Updated last year
- reddit search tool using the pushift.io API☆14Sep 17, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ASCL: adpative Soft Contrastive Learning (ICPR2022)☆22Mar 22, 2025Updated last year
- Paper: https://arxiv.org/abs/1702.02285☆64Dec 19, 2018Updated 7 years ago
- ☆27Nov 2, 2022Updated 3 years ago
- ☆27Feb 8, 2025Updated last year
- Deep Learning - one shot learning for speaker recognition using Filter Banks☆170Jun 23, 2024Updated 2 years ago
- [IJCAI2022] Unsupervised Voice-Face Representation Learning by Cross-Modal Prototype Contrast☆22Oct 25, 2023Updated 2 years ago
- Pytorch implementation of RawNeXt: Speaker verification system for variable-duration utterance with deep layer aggregation and dynamic sc…☆25Jun 22, 2022Updated 4 years ago
- Getting GPU Util 99%☆33Feb 1, 2021Updated 5 years ago
- A lexical normalizer for historical spelling variants using a transformer architecture.☆11Mar 12, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Hypercomplex Neural Networks with PyTorch☆57Jun 11, 2026Updated 2 months ago
- An ambiguous subtitles dataset for visual scene-aware machine translation☆14Oct 17, 2022Updated 3 years ago
- Custom voice prompt packages for Xiaomi\Roborock vacuums☆13Jan 28, 2020Updated 6 years ago
- TensorFlow Lite example on a Raspberry Pi Zero W☆10Dec 16, 2020Updated 5 years ago
- [ECCV2024] The official implementation of "Listen to Look into the Future: Audio-Visual Egocentric Gaze Anticipation".☆17Feb 24, 2025Updated last year
- ☆13Dec 18, 2024Updated last year
- The Unreliability of Explanations in Few-shot Prompting for Textual Reasoning (NeurIPS 2022)☆16Feb 11, 2023Updated 3 years ago