acnn for text-independent speaker recognition
☆10Feb 8, 2022Updated 4 years ago
Alternatives and similar repositories for acnn_speaker_recog
Users that are interested in acnn_speaker_recog are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- TDY-CNN for text-independent speaker verification☆19Nov 7, 2022Updated 3 years ago
- ☆68Sep 13, 2024Updated 2 years ago
- NU-Wave: A Diffusion Probabilistic Model for Neural Audio Upsampling☆37May 25, 2021Updated 5 years ago
- PyTorch implementation of TinyWASE described in our paper "Compressing Speaker Extraction Model with Ultra-low Precision Quantization and…☆11Jun 28, 2021Updated 5 years ago
- SANE-TTS: Stable And Natural End-to-End Multilingual Text-to-Speech☆12Jun 30, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Speechflow for emotion recognition related information decomposition☆10Jul 27, 2021Updated 5 years ago
- ☆10Dec 22, 2023Updated 2 years ago
- Unofficial Pytorch Implementation of WaveGrad2☆111Aug 18, 2021Updated 5 years ago
- Torch implementation of Whisper-guided DDPM based Voice Conversion☆49Mar 7, 2023Updated 3 years ago
- ☆14Sep 20, 2023Updated 3 years ago
- [ICASSP'23] Online speaker clustering☆19Feb 22, 2026Updated 7 months ago
- PyTorch based speaker embedding model☆16Apr 13, 2024Updated 2 years ago
- ☆13Oct 27, 2021Updated 4 years ago
- Predicting Political Instability and Social Conflicts Using Multimodal Data☆10Jun 6, 2016Updated 10 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A SPMI Lab toolkit for language models.☆11Apr 12, 2017Updated 9 years ago
- NMT based punctuation prediction system using lexical and acoustic features .☆14Mar 30, 2020Updated 6 years ago
- Adaptive Convolutions with Per-pixel Dynamic Filter Atom☆27Sep 3, 2021Updated 5 years ago
- Simple torch.nn.module implementation of Alias-Free-GAN style filter and resample☆103Jul 26, 2022Updated 4 years ago
- 将normalize过的中文文本,做逆向normalize。具体功能即实现 chinese_text_normalization的逆向版本。☆13Apr 7, 2021Updated 5 years ago
- This repo gives the code for the official implementation of RCT.☆13Jun 28, 2022Updated 4 years ago
- Implementing VGGVox for Speaker Identification on VoxCeleb1 dataset in PyTorch.☆25Oct 15, 2020Updated 5 years ago
- Simple sinc interpolation in PyTorch.☆15Jul 8, 2023Updated 3 years ago
- Toward Multi Modality Language Model - implementation of GPT-4o/Project Astra☆16Dec 10, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Open source cross-platform implementation of MRCP protocol☆20Mar 3, 2022Updated 4 years ago
- Collection of self-supervised models for speaker and language recognition tasks.☆19Jan 18, 2022Updated 4 years ago
- ☆15May 8, 2021Updated 5 years ago
- Code for the paper: "Leveraging speaker attribute information using multi task learning for speaker verification and diarization" present…☆26Oct 5, 2022Updated 4 years ago
- Pytorch implementation of RNN, CNN, BiGRU and LSTM for text classifcation☆10Apr 30, 2021Updated 5 years ago
- Implementation of the Rhythm Formant Analysis methodology for identifying speech rhythms and rhythm variation in the low frequency spectr…☆17Apr 27, 2023Updated 3 years ago
- wake word spotting with kaldi☆19Dec 3, 2020Updated 5 years ago
- ☆16Jun 13, 2022Updated 4 years ago
- ☆13Oct 24, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Code for Latent Speech-Text Transformer (LST)☆36Mar 12, 2026Updated 6 months ago
- 完整基于omlsa.m实现☆14Nov 26, 2021Updated 4 years ago
- PHO-LID: A Unified Model to Incorporate Acoustic-Phonetic and Phonotactic Information for Language Identification☆21Aug 24, 2023Updated 3 years ago
- ☆35Apr 8, 2019Updated 7 years ago
- The Official Implementation of “Content-Dependent Fine-Grained Speaker Embedding for Zero-Shot Speaker Adaptation in Text-to-Speech Synth…☆86Dec 20, 2022Updated 3 years ago
- [ICML 2023] "Data Efficient Neural Scaling Law via Model Reusing" by Peihao Wang, Rameswar Panda, Zhangyang Wang☆14Jan 4, 2024Updated 2 years ago
- flask+tornado based NVIDIA tacotron2+waveglow tts web app☆28May 25, 2023Updated 3 years ago