A deep learning model for classifying audio frames into [SPEECH, KCHI, CHI, MAL, FEM] classes.
☆52Nov 22, 2025Updated 9 months ago
Alternatives and similar repositories for voice-type-classifier
Users that are interested in voice-type-classifier are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Automatic LInguistic Unit Count Estimator (ALICE)☆51Sep 8, 2026Updated last week
- ACLEW Diarization Virtual Machine☆34Jul 29, 2019Updated 7 years ago
- FEERCI: A Package for Fast non-parametric confidence intervals for Equal Error Rates☆12Mar 13, 2024Updated 2 years ago
- Python package for the management of day-long recordings of children.☆16Jun 29, 2026Updated 2 months ago
- A Whisper to TextGrid script that I use to automatize Corpus Annotation on Praat, with speaker diarization.☆14Nov 14, 2023Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- The Additive Margin SincNet (AM-SincNet) is a new approach for speaker recognition problems which is based in the neural network architec…☆46Oct 3, 2023Updated 2 years ago
- Materials for LOT School 2023, "Language Learning: A Data-Driven Approach"☆14Aug 14, 2024Updated 2 years ago
- Code for the Paper Speech Recognition and Multi-Speaker Diarization of Long Conversations☆39Jun 12, 2023Updated 3 years ago
- Example codes for Audio Processing with Deep Learning & Keras || Presentation ->☆18Jul 16, 2019Updated 7 years ago
- 🎹 pyannote + 🗒 notebook = pyannotebook☆27Jun 12, 2023Updated 3 years ago
- EEGLAB Plug-In for automatic artifact rejection☆18Oct 4, 2017Updated 8 years ago
- Companion repository for the paper "A Comparison of Metric Learning Loss Functions for End-to-End Speaker Verification" published at SLSP…☆61Oct 7, 2020Updated 5 years ago
- Logical inference system based on event semantics and degree semantics in formal semantics☆10Jan 22, 2023Updated 3 years ago
- The Additive Margin MobileNet1D is a new light weight deep learning model for Speaker Recognition which is based on the MobileNetV2 archi…☆31Oct 3, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The PyGpPhs package is designed for utilizing Gaussian Process for port-Hamiltonian system.☆11May 6, 2024Updated 2 years ago
- Automatic subordinate clause extractor☆11Jul 7, 2022Updated 4 years ago
- Example python scripts to evaluate various ASR methods☆11Dec 22, 2021Updated 4 years ago
- WavEncoder is a Python library for encoding audio signals, transforms for audio augmentation, and training audio classification models wi…☆92Jun 6, 2021Updated 5 years ago
- Code and content references for TechnovativeThinker blog posts☆12Jul 6, 2023Updated 3 years ago
- Online streaming speaker change detection model in Pytorch☆44Apr 14, 2023Updated 3 years ago
- Quaternion Neural Networks for 3D Sound Source Localization in Reverberant Environments.☆19Nov 21, 2022Updated 3 years ago
- Tools to prepare linguistic surveys for Amazon Mechanical Turk☆16Oct 3, 2017Updated 8 years ago
- Meta-embeddings are a probabilistic generalization of embeddings in machine learning.☆23Nov 23, 2018Updated 7 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Python package for combining diarization system outputs.☆94Aug 3, 2026Updated last month
- 책 읽어주는 딥러닝을 보고 나도 만들고 싶어져서 공부하며 만드는 repository입니다.☆10Dec 8, 2022Updated 3 years ago
- Accompanies the paper "Learnability and Semantic Universals" ; trains recurrent neural networks to learn to verify sentences with quantif…☆11Aug 10, 2019Updated 7 years ago
- Classify the emotions from variable-length speech segments☆11Mar 29, 2018Updated 8 years ago
- Pytorch implementation of the paper : A Global-local Attention Framework for Weakly Labelled Audio Tagging.☆13Feb 6, 2021Updated 5 years ago
- Automatically exported from code.google.com/p/incremental-top-down-parser☆15Mar 15, 2015Updated 11 years ago
- This is the documentation repository for ongoing releases to the NDA collection 3165. Please add GitHub Issues here for requests or feed…☆17Oct 15, 2025Updated 11 months ago
- For accessing to the dataset, please send your short bio and objective of the study to Dr.Theerawit Wilaiprasitporn (theerawit dot w at v…☆14Apr 29, 2021Updated 5 years ago
- Permutation invariant training in PyTorch☆13Oct 2, 2020Updated 5 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆15Jul 1, 2020Updated 6 years ago
- ☆11Mar 12, 2019Updated 7 years ago
- Tools for calculating psycholinguistically-relevant metrics of language statistics using transformer language models☆13Nov 11, 2022Updated 3 years ago
- Running Mozilla's implementation of Baidu DeepSpeech on Google Colaboratory☆16Mar 18, 2019Updated 7 years ago
- ☆55Oct 17, 2023Updated 2 years ago
- ☆12Jun 22, 2024Updated 2 years ago
- Wav2kws is keyword spotting (KWS) based on Wav2Vec 2.0. This model shows state-of-the-art in Google Speech Commands datasets V1 and V2.☆13Jun 11, 2021Updated 5 years ago