Evaluation script for VoxMovies dataset in PyTorch
☆23Jan 12, 2024Updated 2 years ago
Alternatives and similar repositories for VoxMovies
Users that are interested in VoxMovies are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- VoxSRC2022 workshop development kit☆19Jul 21, 2022Updated 4 years ago
- Audio-visual diarization pipeline used for creating VoxConverse dataset☆22Jun 6, 2025Updated last year
- Simple diarization model☆53Jun 13, 2025Updated last year
- Anonymous ICLR Submission☆14Sep 25, 2019Updated 6 years ago
- ☆42Jan 22, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A toolkit for researchers in the multimodal sound separation.☆16Oct 20, 2023Updated 2 years ago
- PyTorch implementation of Continuous Speech Separation☆12Oct 5, 2022Updated 3 years ago
- repo for active speaker detection for media videos.☆31Nov 19, 2023Updated 2 years ago
- PyTorch implementation of TinyWASE described in our paper "Compressing Speaker Extraction Model with Ultra-low Precision Quantization and…☆11Jun 28, 2021Updated 5 years ago
- Pytorch implementation of Meta-Learning for Short Utterance Speaker Recognition with Imbalance Length Pairs (Interspeech, 2020)☆73Sep 16, 2020Updated 5 years ago
- Unsupervised speech activity detection system.☆11Jul 2, 2018Updated 8 years ago
- Look Who’s Talking: Active Speaker Detection in the Wild☆76Aug 24, 2023Updated 3 years ago
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challenge☆16Mar 26, 2022Updated 4 years ago
- Development Toolkit for the VoxCeleb Speaker Recognition Challenge 2020☆43Jul 17, 2020Updated 6 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Interspeech Tutorial - Resource Efficient and Cross-Modal Learning Toward Foundation Modeling☆15Oct 9, 2023Updated 2 years ago
- ☆18Oct 16, 2018Updated 7 years ago
- Dynamic vision-guided speaker embedding for audio-visual speaker diarization☆12Jul 5, 2022Updated 4 years ago
- ☆16Feb 19, 2026Updated 6 months ago
- ☆12Mar 11, 2025Updated last year
- [CVPR'22 Oral] Temporal Alignment Networks for Long-term Video. Tengda Han, Weidi Xie, Andrew Zisserman.☆122Oct 9, 2023Updated 2 years ago
- ☆23May 11, 2026Updated 3 months ago
- [NeurIPS'22] ReCo: Retrieve and Co-segment for Zero-shot Transfer☆62Apr 20, 2023Updated 3 years ago
- Directional sparse filtering for blind speech separation☆11Jun 8, 2021Updated 5 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆32Jan 6, 2022Updated 4 years ago
- Augmentation adversarial training for self-supervised speaker recognition☆77Aug 15, 2021Updated 5 years ago
- wake-up word emotion recognition [APSIPA 2022]☆17Nov 11, 2022Updated 3 years ago
- Temporal Compact Bilinear Pooling (TCBP)☆11May 27, 2020Updated 6 years ago
- Splits for epic-sounds dataset☆85Aug 2, 2025Updated last year
- 针对CN-Celeb数据集的基于ECAPA-TDNN的说话人识别的pytorch实现☆13Apr 3, 2023Updated 3 years ago
- ☆10Dec 8, 2022Updated 3 years ago
- 机器学习资源☆16May 12, 2020Updated 6 years ago
- [CVPR'22] Official Implementation of "CNLL: A Semi-supervised Approach for Continual Noisy Label Learning"☆18Oct 8, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- In this repository, I try to combine k2 with speechbrain to decode well and fastly.☆16Jun 17, 2022Updated 4 years ago
- Pytorch implementation of Extended U-Net for Speaker Verification in Noisy Environments☆28Jul 24, 2023Updated 3 years ago
- Official implementation of the Odyssey paper "A Probabilistic Fusion Framework for Spoofing Aware Speaker Verification"☆18Jun 24, 2022Updated 4 years ago
- Code for running experiments in our ISMIR'21 paper titled: "Is Disentanglement enough? On Latent Representations for Controllable Music G…☆12Aug 7, 2021Updated 5 years ago
- This repository is the official implementation of our paper "Improving Generalization for AI-Synthesized Voice Detection", which has been…☆24Jan 13, 2026Updated 7 months ago
- ☆26Aug 17, 2026Updated 3 weeks ago
- Implementation of "Look, Listen and Recognise:character-aware audio-visual subtitling"☆21Nov 3, 2025Updated 10 months ago