VoxSRC2022 workshop development kit
☆19Jul 21, 2022Updated 4 years ago
Alternatives and similar repositories for VoxSRC2022
Users that are interested in VoxSRC2022 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Evaluation script for VoxMovies dataset in PyTorch☆23Jan 12, 2024Updated 2 years ago
- Augmentation adversarial training for self-supervised speaker recognition☆77Aug 15, 2021Updated 5 years ago
- Look Who’s Talking: Active Speaker Detection in the Wild☆76Aug 24, 2023Updated 2 years ago
- ☆26Updated this week
- ☆55Oct 17, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Unsupervised Multi-object Segmentation by Predicting Probable Motion Patterns☆17Nov 15, 2022Updated 3 years ago
- [ICASSP'24] Emphasized Non-Target Speaker Knowledge in Knowledge Distillation for Speaker Verification☆16Mar 20, 2024Updated 2 years ago
- Development Toolkit for the VoxCeleb Speaker Recognition Challenge 2020☆43Jul 17, 2020Updated 6 years ago
- Guess What Moves: Unsupervised Video and Image Segmentation by Anticipating Motion☆25Mar 16, 2023Updated 3 years ago
- Development Toolkit for the VoxCeleb Speaker Recognition Challenge 2021☆19Jul 21, 2021Updated 5 years ago
- ICASSP 2023: 'Speaker recognition with two-step multi-modal deep cleansing'☆44Oct 31, 2022Updated 3 years ago
- ☆31Mar 2, 2021Updated 5 years ago
- ☆16Jun 13, 2022Updated 4 years ago
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Simple diarization model☆53Jun 13, 2025Updated last year
- ☆17Jun 30, 2020Updated 6 years ago
- Official Implementation and Dataset of paper - DFADD: The Diffusion and Flow-matching based Audio Deepfake Dataset☆16Apr 7, 2025Updated last year
- Official PyTorch implementation of the paper "Robust Training for Speaker Verification against Noisy Labels" in INTERSPEECH 2023.☆12Oct 23, 2023Updated 2 years ago
- Estimating the Age, Height, and Gender of a speaker with their speech signal.☆15Sep 19, 2022Updated 3 years ago
- Layer-wise analysis of self-supervised pre-trained speech representations☆135Oct 18, 2024Updated last year
- Spot the conversation: speaker diarisation in the wild☆171Jul 26, 2022Updated 4 years ago
- VoxSRC Challenge☆31Jun 11, 2019Updated 7 years ago
- A survey of spoken dialogue models (SDMs) with speech input and speech output. Focus on their Intermediate Representation and Generation …☆32Mar 24, 2026Updated 4 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- CDER (Conversational Diarization Error Rate) Scoring Tool☆22Sep 13, 2022Updated 3 years ago
- ☆51Nov 24, 2022Updated 3 years ago
- A toolkit dedicate for speech evaluation.☆23Sep 26, 2024Updated last year
- ☆23May 11, 2026Updated 3 months ago
- CLASP: Contrastive Language-Speech Pretraining for Multilingual Multimodal Information Retrieval☆13Jun 27, 2025Updated last year
- TMT: Tri-Modal Translation between Speech, Image, and Text by Processing Different Modalities as Different Languages☆18May 23, 2024Updated 2 years ago
- The implementation for "Empowering Whisper as a Joint Multi-Talker and Target-Talker Speech Recognition System".☆34Aug 2, 2025Updated last year
- Highlight code in python multiline strings☆11Oct 10, 2023Updated 2 years ago
- ☆160Jan 9, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Learning Domain-Invariant Transformation for Speaker Verification.☆11Jun 13, 2023Updated 3 years ago
- The official repository of Dynamic-SUPERB.☆200Jun 24, 2025Updated last year
- Official repository for RawNet, RawNet2, and RawNet3☆406Mar 21, 2024Updated 2 years ago
- A simple command line tool to calculate WER for ASR.☆14Jul 28, 2026Updated 3 weeks ago
- Anonymous ICLR Submission☆14Sep 25, 2019Updated 6 years ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- Forced alignment decoder for Whisper.☆16Mar 13, 2024Updated 2 years ago