An in-depth analysis of audio classification on the RAVDESS dataset. Feature engineering, hyperparameter optimization, model evaluation, and cross-validation with a variety of ML techniques and MLP
☆79Nov 5, 2020Updated 5 years ago
Alternatives and similar repositories for sklearn-audio-classification
Users that are interested in sklearn-audio-classification are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Using spectrograms and convolutional neural networks to listen to environment sounds.☆32Jul 23, 2021Updated 5 years ago
- A new comprehensive and diverse few-shot acoustic classification benchmark.☆66Sep 22, 2024Updated last year
- UrbanSound classification using Convolutional Recurrent Networks in PyTorch☆391Jun 16, 2021Updated 5 years ago
- Korean ASR using PyTorch / Listen, Attend and Spell (LAS) / Seq2seq with Attention / Naver-A.I-Hackathon-Speech / A.I Hub Dataset / 한국…☆12Feb 10, 2020Updated 6 years ago
- Urban sounds classification with Covnolutional Neural Networks☆37Nov 15, 2019Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- This repository contains a short introduction on the topic of audio and speech processing -- from basics to applications.☆19Dec 20, 2023Updated 2 years ago
- 基于CNN的音频识别☆18Feb 13, 2019Updated 7 years ago
- Audio classification with VGGish as feature extractor in TensorFlow☆131Dec 4, 2021Updated 4 years ago
- This paper has been accepted in ACM ICMR 2021.☆20Nov 17, 2025Updated 9 months ago
- Detect emotion from audio signals of IEMOCAP dataset using multi-modal approach. Utilized acoustic features, mel-spectrogram and text as …☆41Mar 7, 2024Updated 2 years ago
- small experimentation about positional encoding☆20Feb 9, 2020Updated 6 years ago
- Multi class audio classification using Deep Learning (MLP, CNN): The objective of this project is to build a multi class classifier to id…☆69Jan 8, 2021Updated 5 years ago
- 基于Tensorflow实现声音分类,博客地址:☆106May 8, 2020Updated 6 years ago
- A speech activity detector using HMMs☆11Feb 11, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code for YouTube series: Deep Learning for Audio Classification☆587Feb 6, 2023Updated 3 years ago
- Toolkit to asses speech impairments in patients with neurological disorders☆60May 25, 2018Updated 8 years ago
- Audio classification is a popular topic, here I implement several models using TenserFlow and Keras.☆24Sep 27, 2020Updated 5 years ago
- Python-based cross-platform tool for mining text data (html, transcript, problems) of edX MOOCs on a user's dashboard. It is an extension…☆10Feb 12, 2020Updated 6 years ago
- This is the PyNN code used in the paper titled "Multilayer Spiking Neural Network for audio samples classification using SpiNNaker", whic…☆32Dec 7, 2021Updated 4 years ago
- Simple, straight-forward extraction of acoustic and prosodic features from sound waves based on Praat and Parselmouth.☆29Oct 10, 2019Updated 6 years ago
- Emotive Speech generation based on DAVID: An open-source platform for real-time emotional speech transformation using pysox☆13Feb 20, 2018Updated 8 years ago
- music genre classification : LSTM vs Transformer☆62Mar 25, 2023Updated 3 years ago
- Implementation of the paper "SNR-Based Progressive Learning of Deep Neural Network for Speech Enhancement."☆44Apr 16, 2019Updated 7 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Human emotions are one of the strongest ways of communication. Even if a person doesn’t understand a language, he or she can very well u…☆25Jun 23, 2021Updated 5 years ago
- ☆15May 28, 2020Updated 6 years ago
- compare training duration of CNN with CPU (i7 8550U) vs GPU (mx150) with CUDA depending on batch size☆12Mar 24, 2018Updated 8 years ago
- Word Error Rate Estimation☆16Aug 25, 2020Updated 5 years ago
- Curated List of NLP tutorials☆30Feb 27, 2025Updated last year
- ☆102Apr 1, 2024Updated 2 years ago
- [TMM2022] Source codes of CENet☆42Mar 14, 2023Updated 3 years ago
- Open Source Crimean Tatar Text-to-Speech datasets☆14Feb 23, 2025Updated last year
- Personal playground for AI implemented from scratch in Rust and visualized with WASM. For research and educational purpose.☆10Apr 1, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Predicting various emotion in human speech signal by detecting different speech components affected by human emotion.☆49Aug 2, 2024Updated 2 years ago
- A Python package of the dynamic compressive gammachirp filterbank (dcGC-FB)☆32May 14, 2024Updated 2 years ago
- This repository consists of the IPython Notebook for the work related to audio processing and implementing convolution neural networks fo…☆13Feb 13, 2019Updated 7 years ago
- The details that matter: Frequency resolution of spectrograms in acoustic scene classification - paper replication data☆39Dec 30, 2017Updated 8 years ago
- Speech Emotion Recognition (SER) in Tensorflow using CNNs and CRNNs Based on Mel Spectrograms and Mel Frequency Cepstral Coefficients (MF…☆12Apr 28, 2025Updated last year
- Code for our paper "Efficient Speech Emotion Recognition Using Multi-Scale CNN and Attention" (ICASSP 2021, co-first authorship)☆28Jun 8, 2021Updated 5 years ago
- Implements the blood brain barrier score described in: J. Med. Chem. 2019, 62, 21, 9824-9836 (https://doi.org/10.1021/acs.jmedchem.9b0122…☆14Sep 27, 2020Updated 5 years ago