[INTERSPEECH 2024] Official pytorch code for the paper "Disentangled Representation Learning for Environment-agnostic Speaker Recognition"
☆19Jul 23, 2024Updated 2 years ago
Alternatives and similar repositories for voxceleb-disentangler
Users that are interested in voxceleb-disentangler are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment☆14Feb 5, 2025Updated last year
- [INTERSPEECH 2025] Official code for "SEED: Speaker Embedding Enhancement Diffusion Model"☆61Nov 3, 2025Updated 10 months ago
- Models and codes for INTERSPEECH 2023 paper DistilXLSR: A Light Weight Cross-Lingual Speech Representation Model☆13Mar 30, 2025Updated last year
- ☆10Dec 22, 2023Updated 2 years ago
- INTERSPEECH2023: Target Active Speaker Detection with Audio-visual Cues☆61May 29, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆18Aug 18, 2026Updated 3 weeks ago
- [NeurIPS 2025] AVCD: Mitigating Hallucinations in Audio-Visual Large Language Models through Contrastive Decoding☆27Nov 3, 2025Updated 10 months ago
- Demo for DART, Audio Imagination workshop submission in NeurIPS 2024☆16Apr 22, 2026Updated 4 months ago
- Multi-Stage Face-Voice Association Learning with Keynote Speaker Diarization (ACM MM 2024)☆22Jul 25, 2024Updated 2 years ago
- ☆15Oct 25, 2024Updated last year
- ☆12Jun 14, 2024Updated 2 years ago
- The implementation of "End-to-End Neural Speaker Diarization with an Iterative Adaptive Attractor Estimation", which is accepted by Neura…☆11Aug 27, 2023Updated 3 years ago
- ☆11Sep 4, 2023Updated 3 years ago
- ☆17Oct 24, 2025Updated 10 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆74Feb 15, 2021Updated 5 years ago
- SpeechNAS-Better-Trade-off-between-Latency-and-Accuracy-for-Large-Scale-Speaker-Verification☆30Mar 24, 2023Updated 3 years ago
- Audio-visual diarization pipeline used for creating VoxConverse dataset☆22Jun 6, 2025Updated last year
- ☆37Jan 6, 2026Updated 8 months ago
- This repository contains all the code necessary for running the multilingual distilwhisper from Ferraz et al. 2024 IEEE ICASSP paper.☆34Apr 22, 2026Updated 4 months ago
- Error correction back-end for speaker diarization☆18Sep 26, 2023Updated 2 years ago
- ☆18Jul 22, 2024Updated 2 years ago
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.☆13Feb 13, 2021Updated 5 years ago
- Code and data repository for paper "VoxCeleb enrichment for Age and Gender recognition" submitted at ASRU 2021☆73Dec 18, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Whisper Speech Quality Assessment (WhiSQA)☆16Apr 14, 2026Updated 4 months ago
- ☆11Jul 16, 2026Updated last month
- ☆18Sep 19, 2023Updated 2 years ago
- Lung Extraction from Chest X-ray for Efficient Computing☆15May 13, 2019Updated 7 years ago
- Official implementation of TalkNCE (ICASSP 2024).☆19Apr 30, 2025Updated last year
- Scripts for data generation, scoring and data manifest preparation for CHiME-8 DASR task.☆27Aug 10, 2026Updated last month
- Companion repo for the paper "PixIT: Joint Training of Speaker Diarization and Speech Separation from Real-world Multi-speaker Recordings…☆107Jan 10, 2025Updated last year
- ☆28Dec 22, 2021Updated 4 years ago
- Sound Event Detection (SED) paper collection☆15Jun 26, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆13Oct 27, 2021Updated 4 years ago
- TensorFlow implementation of 'Residual Dense Network for Image Super-Resolution'☆13Mar 28, 2019Updated 7 years ago
- ☆16Feb 19, 2026Updated 6 months ago
- Implementation of "Audio xLSTMs: Learning Self-supervised audio representations with xLSTMs" in PyTorch☆20Aug 28, 2026Updated last week
- ☆18Mar 13, 2024Updated 2 years ago
- ICASSP 2023: 'Speaker recognition with two-step multi-modal deep cleansing'☆44Oct 31, 2022Updated 3 years ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year