This repository contains the official implementation and pretrained weights for the paper "ReDimNet2: Scaling Speaker Verification via Time-Pooled Dimension Reshaping".
☆68Jul 30, 2026Updated this week
Alternatives and similar repositories for redimnet2
Users that are interested in redimnet2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆54Mar 28, 2026Updated 4 months ago
- The official pytorch implemention of the Intespeech 2024 paper "Reshape Dimensions Network for Speaker Recognition"☆206Jul 9, 2026Updated 3 weeks ago
- DiariZen Explained: A Tutorial for the Open Source State-of-the-Art Speaker Diarization Pipeline.☆23Apr 24, 2026Updated 3 months ago
- ☆27Jun 10, 2026Updated last month
- 针对CN-Celeb数据集的基于ECAPA-TDNN的说话人识别的pytorch实现☆13Apr 3, 2023Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A toolkit for speaker diarization.☆509May 29, 2026Updated 2 months ago
- SANE-TTS: Stable And Natural End-to-End Multilingual Text-to-Speech☆11Jun 30, 2023Updated 3 years ago
- This is the official repository of ``Scalable Neural Vocoder from Range-Null Space Decomposition'', which is submitted to TPAMI.☆54Oct 11, 2025Updated 9 months ago
- Research and Production Oriented Speaker Verification, Recognition and Diarization Toolkit☆1,371Jul 8, 2026Updated 3 weeks ago
- [ICASSP 2026] Official code for "Measuring Prosody Diversity in Zero-Shot TTS: A New Metric, Benchmark, and Exploration"☆17Apr 16, 2026Updated 3 months ago
- A solution to denoising and separating for two-speaker-mixed noisy speech, using a BSRNN inspired network.☆15Aug 22, 2023Updated 2 years ago
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequ…☆31Sep 20, 2025Updated 10 months ago
- C++ version of pyannote audio overlapped speech detection pipeline☆13Feb 14, 2024Updated 2 years ago
- The baselines of ARC-Challenge-Interspeech2026☆60Dec 1, 2025Updated 8 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- TS-SEP: Joint Diarization and Separation Conditioned on Estimated Speaker Embeddings☆43Oct 27, 2025Updated 9 months ago
- LLaSE: Maximizing Acoustic Preservation for LLaMA based Speech Enhancement☆16Jul 11, 2025Updated last year
- ☆18Apr 9, 2026Updated 3 months ago
- ☆28Jul 17, 2026Updated 2 weeks ago
- Training code and dataset cleasing with Sidon☆173Apr 24, 2026Updated 3 months ago
- ☆27Aug 29, 2025Updated 11 months ago
- ☆28Apr 6, 2026Updated 3 months ago
- [INTERSPEECH 2025] Official code for "SEED: Speaker Embedding Enhancement Diffusion Model"☆59Nov 3, 2025Updated 9 months ago
- The VoxTube dataset official repository☆71Feb 14, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆26Mar 31, 2026Updated 4 months ago
- FINALLY: Fast and universal speech enhancement model delivering studio-quality audio for a wide range of recordings.☆28Apr 1, 2026Updated 4 months ago
- SLT 2024 Mandarin Stuttering Event Detection and Automatic Speech Recognition Challenge☆12Jun 11, 2024Updated 2 years ago
- Pure-PyTorch Parakeet TDT inference☆51Mar 10, 2026Updated 4 months ago
- Official Repository For VoxBlink2☆89Aug 13, 2024Updated last year
- C++ version of pyannote audio speaker diarizaiton pipeline☆22Feb 14, 2024Updated 2 years ago
- Attention-Based Encoder-Decoder Target-Speaker Voice Activity Detection for Robust Speaker Diarization☆31Sep 22, 2025Updated 10 months ago
- Final training script from HuggingFace Whisper Fine tuning event - to get best results on finetuned model.☆12Dec 24, 2022Updated 3 years ago
- PyTorch implementation of Miipher-2 [2025] which is a speech restoration model by Google DeepMind☆70Sep 22, 2025Updated 10 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official repository of UniPASE, a SOTA USE model☆54Jul 21, 2026Updated last week
- ESLTTS dataset☆16Feb 6, 2025Updated last year
- ASLP Summer Inter@NPU☆13Jul 30, 2024Updated 2 years ago
- ☆68Aug 16, 2023Updated 2 years ago
- [ICLR2026] FlexiCodec: A Dynamic Neural Audio Codec for Low Frame Rates☆51Jul 1, 2026Updated last month
- Multi-Stage Face-Voice Association Learning with Keynote Speaker Diarization (ACM MM 2024)☆22Jul 25, 2024Updated 2 years ago
- Pushing the Limits of Zero-shot End-to-End Speech Translation☆25Dec 12, 2024Updated last year