A deep learning model for classifying audio frames into [SPEECH, KCHI, CHI, MAL, FEM] classes.
☆51Nov 22, 2025Updated 9 months ago
Alternatives and similar repositories for voice-type-classifier
Users that are interested in voice-type-classifier are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ACLEW Diarization Virtual Machine☆34Jul 29, 2019Updated 7 years ago
- Behavioral probing of language acquisition models at the lexical and syntactic level☆20Jul 17, 2023Updated 3 years ago
- Python package for the management of day-long recordings of children.☆16Jun 29, 2026Updated 2 months ago
- The Additive Margin SincNet (AM-SincNet) is a new approach for speaker recognition problems which is based in the neural network architec…☆46Oct 3, 2023Updated 2 years ago
- Code for the Paper Speech Recognition and Multi-Speaker Diarization of Long Conversations☆39Jun 12, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Example codes for Audio Processing with Deep Learning & Keras || Presentation ->☆18Jul 16, 2019Updated 7 years ago
- 🎹 pyannote + 🗒 notebook = pyannotebook☆27Jun 12, 2023Updated 3 years ago
- EEGLAB Plug-In for automatic artifact rejection☆18Oct 4, 2017Updated 8 years ago
- Da - ECHO - RetrievAl - daTasEt☆37Jul 7, 2024Updated 2 years ago
- Companion repository for the paper "A Comparison of Metric Learning Loss Functions for End-to-End Speaker Verification" published at SLSP…☆61Oct 7, 2020Updated 5 years ago
- The Additive Margin MobileNet1D is a new light weight deep learning model for Speaker Recognition which is based on the MobileNetV2 archi…☆31Oct 3, 2023Updated 2 years ago
- Example python scripts to evaluate various ASR methods☆11Dec 22, 2021Updated 4 years ago
- Online streaming speaker change detection model in Pytorch☆44Apr 14, 2023Updated 3 years ago
- a close enough approximation of the shadertoy framework☆12Jul 2, 2020Updated 6 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Meta-embeddings are a probabilistic generalization of embeddings in machine learning.☆23Nov 23, 2018Updated 7 years ago
- Tutorial session material of Pytest in PyCon KR 2019☆10Jul 22, 2026Updated last month
- A simple pyaudio microphone interface☆11Jul 27, 2018Updated 8 years ago
- Permutation invariant training in PyTorch☆13Oct 2, 2020Updated 5 years ago
- Self-Supervised Speech/Sound Pre-training and Representation Learning Toolkit☆13Nov 18, 2022Updated 3 years ago
- ☆11Mar 12, 2019Updated 7 years ago
- Running Mozilla's implementation of Baidu DeepSpeech on Google Colaboratory☆16Mar 18, 2019Updated 7 years ago
- ☆55Oct 17, 2023Updated 2 years ago
- ☆13Jan 10, 2017Updated 9 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Wav2kws is keyword spotting (KWS) based on Wav2Vec 2.0. This model shows state-of-the-art in Google Speech Commands datasets V1 and V2.☆13Jun 11, 2021Updated 5 years ago
- An attempt to replicate the results of [1706.08612] VoxCeleb: a large-scale speaker identification dataset☆12Dec 11, 2019Updated 6 years ago
- 2019 PyCon kr tutorial: "네이버 영화 평점 데이터로 자연어처리 논문 구현 시작하기"☆13Aug 21, 2019Updated 7 years ago
- Implementation and Deployment of Multilingual Custom Keyword Spotting Running in Real-time on an Edge Device.☆11Apr 27, 2023Updated 3 years ago
- The Batch Electroencephalography Automated Processing Platform (BEAPP)☆40Sep 15, 2025Updated 11 months ago
- Learnable STRF, from Riad et al. 2021 JASA☆13Aug 21, 2021Updated 5 years ago
- R labs for the book OpenIntro Statistics (https://www.openintro.org/stat/)☆13Nov 17, 2016Updated 9 years ago
- 친절한 실전 딥러닝 수업☆12Sep 22, 2020Updated 5 years ago
- Flexible, extensible and scalable web-based speech annotation tool☆14Apr 4, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- My system for the DCASE 2022 Task 3 Sound Event Localizaiton and Detection.☆12Nov 12, 2022Updated 3 years ago
- ☆22Jun 30, 2021Updated 5 years ago
- ☆70Feb 16, 2017Updated 9 years ago
- Estimating the Age, Height, and Gender of a speaker with their speech signal.☆15Sep 19, 2022Updated 3 years ago
- VoxSRC2022 workshop development kit☆19Jul 21, 2022Updated 4 years ago
- PyTorch based speaker embedding model☆16Apr 13, 2024Updated 2 years ago
- Raw waveform adaptation with SincNet☆12Mar 19, 2024Updated 2 years ago