A ResNet Speaker Recognition&Verification Demo
☆27Oct 19, 2021Updated 4 years ago
Alternatives and similar repositories for Speaker-Recognition-Demo
Users that are interested in Speaker-Recognition-Demo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ICASSP 2022: 'Self-supervised Speaker Recognition with Loss-gated Learning'☆91May 29, 2023Updated 3 years ago
- ACM MM 2021: 'Is Someone Speaking? Exploring Long-term Temporal Features for Audio-visual Active Speaker Detection'☆497Oct 23, 2023Updated 2 years ago
- Unofficial reimplementation of ECAPA-TDNN for speaker recognition (EER=0.86 for Vox1_O when train only in Vox2)☆828Apr 11, 2024Updated 2 years ago
- ☆28Jan 21, 2026Updated 7 months ago
- INTERSPEECH2023: Target Active Speaker Detection with Audio-visual Cues☆61May 29, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- tts fronted-end☆11Dec 19, 2018Updated 7 years ago
- Tools for downloading VoxCeleb2 dataset☆35Mar 16, 2024Updated 2 years ago
- Generate accompaniment part with chords using Evolutionary algorithm.☆11May 8, 2022Updated 4 years ago
- An evolutionary algorithm that generates an accompaniment to a given melody that consists of triad chords while following music theory ru…☆10Sep 19, 2022Updated 3 years ago
- Code for Audio-Visual Target Speaker Extraction with Selective Auditory Attention (TASLP)☆35Feb 28, 2025Updated last year
- Speaker verification task with ECAPA-TDNN model (trained on Persian dataset)☆12Sep 15, 2022Updated 3 years ago
- ☆17Jan 31, 2023Updated 3 years ago
- The official repo/implementation of the paper "Training a Singing Transcription Model Using Connectionist Temporal Classification Loss an…☆13Mar 25, 2025Updated last year
- GSoC'16 RedHen Labs☆11Aug 22, 2016Updated 10 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A Pytorch implementation of 'Progressive Neural Networks for Transfer Learning in Emotion Recognition'☆11Jul 31, 2018Updated 8 years ago
- ☆43Nov 22, 2024Updated last year
- Classify the emotions from variable-length speech segments☆11Mar 29, 2018Updated 8 years ago
- A curated list of awesome speaker recognition/verification papers, projects, datasets, and competition.☆15Aug 29, 2021Updated 5 years ago
- ☆10Jun 2, 2021Updated 5 years ago
- demos, exercise sets, and studios for java-web-development☆11Jun 14, 2023Updated 3 years ago
- An attempt at genre classification with convolutional neural networks and spectrograms☆15Nov 25, 2017Updated 8 years ago
- An attempt to replicate the results of [1706.08612] VoxCeleb: a large-scale speaker identification dataset☆12Dec 11, 2019Updated 6 years ago
- [NeurIPS 2024] SD-Eval: A Benchmark Dataset for Spoken Dialogue Understanding Beyond Words☆57Jun 25, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The code for AIM2022 compressed image super-resolution☆16Apr 27, 2023Updated 3 years ago
- Attention Backend for Aotumatic Speaker Verification with Multiple Enrollment Utterances☆50Oct 27, 2022Updated 3 years ago
- Latest PyTorch Implementation of DeltaGRU & DeltaLSTM that Exploits Temporal Sparsity in Sequential Data☆18Sep 30, 2023Updated 2 years ago
- In defence of metric learning for speaker recognition☆1,175Apr 22, 2026Updated 4 months ago
- Momentum Contrast for Unsupervised Visual Representation Learning☆16Mar 24, 2023Updated 3 years ago
- Cross-Modal Relation-Aware Networks for Audio-Visual Event Localization, ACM MM 2020☆33Nov 6, 2020Updated 5 years ago
- Master's Thesis: Automatic Tagging of Musical Compositions Using Machine Learning Methods☆17May 22, 2023Updated 3 years ago
- Python code for training and testing of GMM-UBM and maximum a posterirori (MAP) adaptation based speaker verification☆20Jul 31, 2020Updated 6 years ago
- Implementation of the VGGVox network using TensorFlow.☆17Mar 20, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- AD-TUNING: An Adaptive CHILD-TUNING Approach to Efficient Hyperparameter Optimization of Child Networks for Speech Processing Tasks in th…☆11Feb 23, 2024Updated 2 years ago
- ☆16Feb 19, 2026Updated 6 months ago
- neural network and loss for asv implemented by PyTorch. (Triplet loss, LMCL, Angular Loss, Softmax)☆21Oct 23, 2019Updated 6 years ago
- Code of paper "Densely Connected Pyramidal Dilated Convolutional Network for Hyperspectral Image Classification"☆10Jun 21, 2022Updated 4 years ago
- 单独维护的中文TTS☆34Oct 28, 2022Updated 3 years ago
- ☆11Nov 5, 2025Updated 9 months ago
- Text independent speaker recognition algorithm based on CNN☆24Aug 30, 2025Updated last year