A speaker gender classifier. MFC feature engineering and a pre-trained ResNet-50. GradCAM interpretation.
☆27Nov 18, 2021Updated 4 years ago
Alternatives and similar repositories for gender-audio-classification
Users that are interested in gender-audio-classification are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Audio classification is a popular topic, here I implement several models using TenserFlow and Keras.☆24Sep 27, 2020Updated 5 years ago
- Machine learning experiment to perform gender classification from raw audio.☆23Sep 1, 2018Updated 7 years ago
- ☆21Mar 8, 2020Updated 6 years ago
- Voice based gender recognition using Mel-frequency cepstrum coefficients (MFCC) and Gaussian mixture models (GMM)☆221Jul 6, 2023Updated 3 years ago
- Classifying 10 different categories of Sound using Deep Learning.☆25Jul 21, 2018Updated 8 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Building a Deep learning model that predicts the gender of a speaker using TensorFlow 2☆130Apr 25, 2023Updated 3 years ago
- This is a repository for the code developed to produced the results in the paper "Detection of Inferior Myocardial Infarction using Shall…☆10May 15, 2024Updated 2 years ago
- Processing ECG Signal, QRS and ST Segment Detection, BPM Calculation, ST Slope Measurement and Myocardial Ischemia Detection.☆12Jun 27, 2020Updated 6 years ago
- Audio classification via transfer learning☆35Oct 3, 2019Updated 6 years ago
- Music generation using Elementary Cellular Automata.☆13Nov 23, 2015Updated 10 years ago
- Audio classification using Machine Learning☆13Dec 17, 2015Updated 10 years ago
- Deep learning application for predicting ocean wave behaviors.☆16May 31, 2020Updated 6 years ago
- cpp inference for EmotiVoice☆16Jan 1, 2024Updated 2 years ago
- NeMo: a toolkit for conversational AI☆13May 4, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official Implementation and Dataset of paper - DFADD: The Diffusion and Flow-matching based Audio Deepfake Dataset☆16Apr 7, 2025Updated last year
- ☆13Nov 12, 2024Updated last year
- ☆10Jul 27, 2021Updated 5 years ago
- CNN-LSTM based QRS detector for ECG signals☆25Apr 7, 2020Updated 6 years ago
- A CNN audio classifier via spectrogram images.☆10Jul 21, 2017Updated 9 years ago
- ☆10Jan 20, 2024Updated 2 years ago
- ☆40Feb 26, 2026Updated 5 months ago
- E-BATS: Efficient Backpropagation-Free Test-Time Adaptation for Speech Foundation Models☆18Nov 4, 2025Updated 8 months ago
- FastLongSpeech is a novel framework designed to extend the capabilities of Large Speech-Language Models for efficient long-speech process…☆16Jul 22, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This repository includes the code to reproduce our paper Partially-Connected Differentiable Architecture Search for Deepfake and Spoofing…☆18Apr 30, 2022Updated 4 years ago
- This network uses Unet to perform segmentation of ECG to identify P,QRS,T components of a given ECG.☆22Jul 23, 2018Updated 8 years ago
- Text classification applications for detecting fraudulent calls, spam messages, etc.☆14Jan 29, 2024Updated 2 years ago
- Pytorch implementation of "A Differentiable Perceptual Audio Metric Learned from Just Noticeable Differences", Pranay Manocha et al. - un…☆65Apr 2, 2020Updated 6 years ago
- ♂️♀️ Detect a person's gender from a voice file (90.7% +/- 1.3% accuracy).☆91Jun 17, 2024Updated 2 years ago
- ☆17Aug 27, 2025Updated 11 months ago
- ☆13Mar 25, 2021Updated 5 years ago
- Room Impulse Response Generator☆19May 20, 2018Updated 8 years ago
- ☆16Nov 19, 2021Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A Tensorflow implementation of Speech Emotion Recognition using Audio signals and Text Data☆12May 16, 2022Updated 4 years ago
- Using LSTMs to classify ECG signals in several different heart diseases. Data from the PTB diagnostic database.☆24Jul 26, 2018Updated 8 years ago
- BERT score for text generation☆12Jan 15, 2025Updated last year
- ☆34Nov 19, 2023Updated 2 years ago
- ☆11May 18, 2022Updated 4 years ago
- Submission to MediaEval 2021 Emotions and Themes in Music challenge. Noisy-student training for music emotion tagging☆11Dec 2, 2021Updated 4 years ago
- ☆10Nov 10, 2021Updated 4 years ago