Machine learning experiment to perform gender classification from raw audio.
☆23Sep 1, 2018Updated 8 years ago
Alternatives and similar repositories for raw-audio-gender-classification
Users that are interested in raw-audio-gender-classification are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A module for normalising text.☆10Nov 6, 2019Updated 6 years ago
- A speaker gender classifier. MFC feature engineering and a pre-trained ResNet-50. GradCAM interpretation.☆27Nov 18, 2021Updated 4 years ago
- A packaged convolutional voice activity detector for noisy environments.☆14Jun 15, 2019Updated 7 years ago
- Two Keras models for child/adult & man/woman classify use speech in Python.☆14Jan 19, 2019Updated 7 years ago
- A speaker embedding network in Pytorch that is very quick to set up and use for whatever purposes.☆91Apr 2, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- Final training script from HuggingFace Whisper Fine tuning event - to get best results on finetuned model.☆12Dec 24, 2022Updated 3 years ago
- DUSTED: Spoken-Term Discovery using Discrete Speech Units☆17Oct 2, 2024Updated last year
- Pytorch implementation of Meta-Learning for Short Utterance Speaker Recognition with Imbalance Length Pairs (Interspeech, 2020)☆73Sep 16, 2020Updated 5 years ago
- ☆11Nov 10, 2015Updated 10 years ago
- ☆12Dec 22, 2020Updated 5 years ago
- Tools and scripts for working with ELAN☆10Aug 4, 2022Updated 4 years ago
- ☆11Apr 6, 2019Updated 7 years ago
- This repository created for the NHN ASR hackathon competition.☆12Sep 20, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- NeMo: a toolkit for conversational AI☆13May 4, 2024Updated 2 years ago
- Code on IART: Intent-aware Response Ranking with Transformers in Information-seeking Conversation Systems (WWW 2020)☆11Apr 18, 2021Updated 5 years ago
- This repository contains the research project that enables the robot to automatically join a group based on the modeled personal, social …☆11Nov 4, 2018Updated 7 years ago
- Converts CLIP models to ONNX☆11Jan 17, 2023Updated 3 years ago
- ☆19Dec 8, 2020Updated 5 years ago
- Supervised Machine Learning Classification Algorithms using Python and R (Logistic Regression, Decision Tree, Random Forest, SVM)☆14May 12, 2018Updated 8 years ago
- Tutorial session material of Pytest in PyCon KR 2019☆10Jul 22, 2026Updated last month
- Using machine learning to recognise gender by analysing recorded voice.☆12Nov 7, 2025Updated 9 months ago
- An implementation of YOLOv1 for object detection in TensorFlow.☆12Sep 29, 2020Updated 5 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- A deep neural network for finding text-independent speaker embedding written in tensorflow and tensorpack☆10Feb 19, 2018Updated 8 years ago
- Paper List for Dialogue and Interactive Systems☆15Jun 5, 2020Updated 6 years ago
- ☆11Mar 12, 2019Updated 7 years ago
- Official code for Tell Me What You See: A Zero-Shot Action Recognition Method Based on Natural Language Descriptions (Multimedia Tools an…☆13Mar 8, 2024Updated 2 years ago
- Gender recognition by voice and speech analysis☆362Jan 16, 2023Updated 3 years ago
- Weakly-supervised action segmentation in video☆16Feb 13, 2022Updated 4 years ago
- Deep learning-based audio spoofing attack detection experiments for speaker verification.☆14Aug 9, 2026Updated 3 weeks ago
- 2019 PyCon kr tutorial: "네이버 영화 평점 데이터로 자연어처리 논문 구현 시작하기"☆13Aug 21, 2019Updated 7 years ago
- A library for loading, modifying and saving BVH motion capture files.☆13Dec 15, 2010Updated 15 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- FEERCI: A Package for Fast non-parametric confidence intervals for Equal Error Rates☆12Mar 13, 2024Updated 2 years ago
- The Additive Margin SincNet (AM-SincNet) is a new approach for speaker recognition problems which is based in the neural network architec…☆46Oct 3, 2023Updated 2 years ago
- PyTorch based speaker embedding model☆16Apr 13, 2024Updated 2 years ago
- [AAAI 2018] Implementation of the Ethics Shaping approach proposed in "A low-cost ethics shaping approach for designing reinforcement lea…☆11Aug 3, 2018Updated 8 years ago
- Attention Backend for Aotumatic Speaker Verification with Multiple Enrollment Utterances☆50Oct 27, 2022Updated 3 years ago
- GUI for albumentations library☆11Sep 13, 2019Updated 6 years ago
- Code for auto-generating maze distractors and running maze in ibex☆25Jul 27, 2026Updated last month