Efficient Speech Processing Tookit for Automatic Speaker Recognition
☆18Feb 8, 2023Updated 3 years ago
Alternatives and similar repositories for sugar
Users that are interested in sugar are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Sequence algorithms for use in Flashlight.☆14Jan 12, 2026Updated 7 months ago
- Docker image and scripts for training finetuned or completely personal Kaldi speech models. Particularly for use with kaldi-active-gramma…☆21Jan 24, 2022Updated 4 years ago
- PAVOQUE Corpus of Expressive Speech☆12Aug 2, 2016Updated 10 years ago
- Official repository of NeXt-TDNN for speaker verification☆86Oct 10, 2024Updated last year
- Predicting Political Instability and Social Conflicts Using Multimodal Data☆10Jun 6, 2016Updated 10 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Simplified recipes for preparing commonly used speech datasets, and a PyTorch-compatible Python data loader that can perform standard fea…☆15Jun 25, 2026Updated last month
- code for paper "learning to fool the speaker recognition"☆10Jun 12, 2020Updated 6 years ago
- Implementation of CoBERT: Self-Supervised Speech Representation Learning Through Code Representation Learning☆48Nov 8, 2023Updated 2 years ago
- Acoustic Echo Cancellation☆14May 29, 2022Updated 4 years ago
- Data from "Crowdsourcing of Parallel Corpora: the Case of Style Transfer for Detoxification" paper☆14Apr 3, 2025Updated last year
- acnn for text-independent speaker recognition☆10Feb 8, 2022Updated 4 years ago
- LightHuBERT: Lightweight and Configurable Speech Representation Learning with Once-for-All Hidden-Unit BERT☆74Sep 26, 2022Updated 3 years ago
- ICASSP 2022: 'Self-supervised Speaker Recognition with Loss-gated Learning'☆91May 29, 2023Updated 3 years ago
- Official Repository For VoxBlink2☆88Aug 13, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Example python scripts to evaluate various ASR methods☆11Dec 22, 2021Updated 4 years ago
- A simple universal data description format for datasets, tailored for interfacing with humans.☆25Feb 16, 2021Updated 5 years ago
- WavEncoder is a Python library for encoding audio signals, transforms for audio augmentation, and training audio classification models wi…☆92Jun 6, 2021Updated 5 years ago
- This repository created for the NHN ASR hackathon competition.☆12Sep 20, 2023Updated 2 years ago
- Pytorch implementation of RNN, CNN, BiGRU and LSTM for text classifcation☆10Apr 30, 2021Updated 5 years ago
- codes for Neural Architecture Ranker and detailed cell information datasets based on NAS-Bench series☆12Jul 11, 2022Updated 4 years ago
- An SSH implemenation in pure Haskell☆17Feb 14, 2022Updated 4 years ago
- Optimized Inference with ResNet-50: A demonstration of inference performance using PyTorch, TensorRT, ONNX, and OpenVINO. Includes benchm…☆10Nov 3, 2025Updated 9 months ago
- Converts TensorFlow checkpoints (with index, meta and data files) to PyTorch, HDF5 and JSON☆18Feb 26, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [InterSpeech 2020] "AutoSpeech: Neural Architecture Search for Speaker Recognition" by Shaojin Ding*, Tianlong Chen*, Xinyu Gong, Weiwei …☆206Dec 8, 2022Updated 3 years ago
- Tutorial session material of Pytest in PyCon KR 2019☆10Jul 22, 2026Updated 3 weeks ago
- 책 읽어주는 딥러닝을 보고 나도 만들고 싶어져서 공부하며 만드는 repository입니다.☆10Dec 8, 2022Updated 3 years ago
- ☆16Feb 19, 2026Updated 5 months ago
- BurrMill core☆22Nov 2, 2021Updated 4 years ago
- Speaker Diarization using GRU in PyTorch☆11Aug 29, 2020Updated 5 years ago
- Unofficial reimplementation of ECAPA-TDNN for speaker recognition (EER=0.86 for Vox1_O when train only in Vox2)☆825Apr 11, 2024Updated 2 years ago
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- Voice100 includes neural TTS/ASR models. Inference of Voice100 is low cost as its models are tiny and only depend on CNN without autoregr…☆28Nov 23, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Python scripts to create noisy and reverberant 2-speaker mixture audio with Libri-Light and WHAM☆17Nov 7, 2024Updated last year
- Experiments with GAN, WGAN, WGAN-GP, DC-GAN, cGAN, AC,GAN and pix2pix☆10May 28, 2019Updated 7 years ago
- A deep neural network for finding text-independent speaker embedding written in tensorflow and tensorpack☆10Feb 19, 2018Updated 8 years ago
- [NeurIPS 2024] SD-Eval: A Benchmark Dataset for Spoken Dialogue Understanding Beyond Words☆57Jun 25, 2024Updated 2 years ago
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.☆13Oct 11, 2022Updated 3 years ago
- SOME/IP guide☆19Jun 5, 2023Updated 3 years ago
- ☆26Nov 19, 2020Updated 5 years ago