Efficient Speech Processing Tookit for Automatic Speaker Recognition
☆18Feb 8, 2023Updated 3 years ago
Alternatives and similar repositories for sugar
Users that are interested in sugar are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- TDY-CNN for text-independent speaker verification☆19Nov 7, 2022Updated 3 years ago
- Sequence algorithms for use in Flashlight.☆14Jan 12, 2026Updated 8 months ago
- Docker image and scripts for training finetuned or completely personal Kaldi speech models. Particularly for use with kaldi-active-gramma…☆21Jan 24, 2022Updated 4 years ago
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.☆13Feb 13, 2021Updated 5 years ago
- PAVOQUE Corpus of Expressive Speech☆12Aug 2, 2016Updated 10 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Pytorch implementation of Meta-Learning for Short Utterance Speaker Recognition with Imbalance Length Pairs (Interspeech, 2020)☆73Sep 16, 2020Updated 6 years ago
- Official repository of NeXt-TDNN for speaker verification☆86Oct 10, 2024Updated last year
- Simplified recipes for preparing commonly used speech datasets, and a PyTorch-compatible Python data loader that can perform standard fea…☆15Jun 25, 2026Updated 2 months ago
- code for paper "learning to fool the speaker recognition"☆10Jun 12, 2020Updated 6 years ago
- Implementation of CoBERT: Self-Supervised Speech Representation Learning Through Code Representation Learning☆48Nov 8, 2023Updated 2 years ago
- Data from "Crowdsourcing of Parallel Corpora: the Case of Style Transfer for Detoxification" paper☆14Apr 3, 2025Updated last year
- acnn for text-independent speaker recognition☆10Feb 8, 2022Updated 4 years ago
- LightHuBERT: Lightweight and Configurable Speech Representation Learning with Once-for-All Hidden-Unit BERT☆74Sep 26, 2022Updated 3 years ago
- ICASSP 2022: 'Self-supervised Speaker Recognition with Loss-gated Learning'☆91May 29, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official Repository For VoxBlink2☆90Aug 13, 2024Updated 2 years ago
- Example python scripts to evaluate various ASR methods☆11Dec 22, 2021Updated 4 years ago
- A simple universal data description format for datasets, tailored for interfacing with humans.☆25Feb 16, 2021Updated 5 years ago
- Haskell + nixpkgs = nix-hs☆24Jun 2, 2021Updated 5 years ago
- Pytorch implementation of RNN, CNN, BiGRU and LSTM for text classifcation☆10Apr 30, 2021Updated 5 years ago
- codes for Neural Architecture Ranker and detailed cell information datasets based on NAS-Bench series☆12Aug 18, 2026Updated last month
- An SSH implemenation in pure Haskell☆17Feb 14, 2022Updated 4 years ago
- Converts TensorFlow checkpoints (with index, meta and data files) to PyTorch, HDF5 and JSON☆18Feb 26, 2021Updated 5 years ago
- [InterSpeech 2020] "AutoSpeech: Neural Architecture Search for Speaker Recognition" by Shaojin Ding*, Tianlong Chen*, Xinyu Gong, Weiwei …☆207Dec 8, 2022Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Tutorial session material of Pytest in PyCon KR 2019☆10Jul 22, 2026Updated 2 months ago
- 책 읽어주는 딥러닝을 보고 나도 만들고 싶어져서 공부하며 만드는 repository입니다.☆10Dec 8, 2022Updated 3 years ago
- ☆16Feb 19, 2026Updated 7 months ago
- BurrMill core☆22Nov 2, 2021Updated 4 years ago
- Speaker Diarization using GRU in PyTorch☆11Aug 29, 2020Updated 6 years ago
- Unofficial reimplementation of ECAPA-TDNN for speaker recognition (EER=0.86 for Vox1_O when train only in Vox2)☆826Apr 11, 2024Updated 2 years ago
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- Voice100 includes neural TTS/ASR models. Inference of Voice100 is low cost as its models are tiny and only depend on CNN without autoregr…☆28Nov 23, 2023Updated 2 years ago
- Python scripts to create noisy and reverberant 2-speaker mixture audio with Libri-Light and WHAM☆17Nov 7, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Experiments with GAN, WGAN, WGAN-GP, DC-GAN, cGAN, AC,GAN and pix2pix☆10May 28, 2019Updated 7 years ago
- A deep neural network for finding text-independent speaker embedding written in tensorflow and tensorpack☆10Feb 19, 2018Updated 8 years ago
- [NeurIPS 2024] SD-Eval: A Benchmark Dataset for Spoken Dialogue Understanding Beyond Words☆57Jun 25, 2024Updated 2 years ago
- The project tries to solve a speaker diarization problem using audio features, face recognition and video feature extraction from face im…☆16Feb 10, 2019Updated 7 years ago
- ☆11Mar 12, 2019Updated 7 years ago
- ☆26Nov 19, 2020Updated 5 years ago
- Code and data repository for paper "VoxCeleb enrichment for Age and Gender recognition" submitted at ASRU 2021☆73Dec 18, 2021Updated 4 years ago