Classification of 11 types of audio clips using MFCCs features and LSTM. Pretrained on Speech Command Dataset with intensive data augmentation.
☆43Dec 14, 2022Updated 3 years ago
Alternatives and similar repositories for Speech-Commands-Classification-by-LSTM-PyTorch
Users that are interested in Speech-Commands-Classification-by-LSTM-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Speech command classification on Speech-Command v0.02 dataset using PyTorch and torchaudio. In this example, three models have been train…☆10Dec 5, 2022Updated 3 years ago
- ☆21Mar 8, 2020Updated 6 years ago
- Speech Commands Recognition using end-to-end deep learning models in pytorch☆28Oct 8, 2020Updated 5 years ago
- Speech command recognition with capsule network & various NNs / KWS on Google Speech Command Dataset.☆25Jan 28, 2019Updated 7 years ago
- A CNN audio classifier via spectrogram images.☆10Jul 21, 2017Updated 9 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- PyTorch implementations of neural network models for keyword spotting☆11Oct 19, 2020Updated 5 years ago
- VoiceCode is an Open Source initiative started by the National Research Council of Canada, to develop a programming by voice toolbox. The…☆10Apr 17, 2020Updated 6 years ago
- Triangle Attack: A Query-efficient Decision-based Adversarial Attack (ECCV 2022)☆16Jul 19, 2022Updated 4 years ago
- Speech commands recognition with PyTorch | Kaggle 10th place solution in TensorFlow Speech Recognition Challenge☆201Jan 19, 2024Updated 2 years ago
- Speech Recognition for speakers with speech disorders due to diseases like Cerebral Palsy, Parkinson or Amyotrophic Lateral Sclerosis ALS…☆23Mar 26, 2017Updated 9 years ago
- Repository for Weak Label Learning for Audio Events - A closer look. Uses Audioset subset data provided for reproducibility.☆32Sep 13, 2023Updated 3 years ago
- A Python 2.7 implementation of Mel Frequency Cepstral Coefficients (MFCC) and Dynamic Time Warping (DTW) algorithms for Automated Speech …☆16Apr 23, 2018Updated 8 years ago
- Classify the emotions from variable-length speech segments☆11Mar 29, 2018Updated 8 years ago
- ☆17Apr 12, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Implementation of a standard synthetic speech detection approach via bob.bio.spear Python package☆16Mar 23, 2017Updated 9 years ago
- Signal processing and clustering of EEG signals☆11Sep 13, 2018Updated 8 years ago
- Google Speech Command Dataset Classification Neural Network, CNN, RNN☆26Aug 29, 2017Updated 9 years ago
- https://dodiku.github.io/audio_noise_clustering/results/ ==> An experiment with a variety of clustering (and clustering-like) techniques …☆26May 5, 2017Updated 9 years ago
- System for Emotion Detection in given speech data using joint modelling of hand crafted prosody rich features , MFCC features and LSTM ba…☆10Nov 15, 2017Updated 8 years ago
- Mispronunciation detection code for jingju singing voice☆19Sep 5, 2018Updated 8 years ago
- Attention-based model for keywords spotting☆19Aug 9, 2021Updated 5 years ago
- numpy implementation of Recurrent Neural Network☆41Jun 2, 2019Updated 7 years ago
- This repository is developed in MATLAB. Speech Augmentation is based on Adaptive Filtering while Endpoint Detection is based on Voice Act…☆10Dec 7, 2020Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Predicting Billboard's Year-End Hot 100 Songs using audio features from Spotify and lyrics from Musixmatch☆18Jul 14, 2024Updated 2 years ago
- A set of speech feature extraction functions for ASR and speaker identification written in matlab.☆43Oct 28, 2016Updated 9 years ago
- Audio feature extraction and multi-classification with the ECS-10 data set☆21Jun 7, 2018Updated 8 years ago
- Classify various radiology images into respective categories.Being done using various shape and texture features for feature extraction a…☆10May 20, 2016Updated 10 years ago
- Recognizing common speech commands using Keras and Tensorflow.☆10Dec 17, 2018Updated 7 years ago
- Flask webapp/endpoint that compares the user's speech with different accents and assigns similarity scores based on speed, voice (DTW/MFC…☆18Jun 27, 2017Updated 9 years ago
- Counts frequencies of words using movie and television subtitles.☆19Jan 26, 2015Updated 11 years ago
- Acoustic and language models for minorised languages.☆26Jul 17, 2026Updated 2 months ago
- A repository of ELL models☆21Jan 16, 2026Updated 8 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Hotword detector wrapper of Snowboy by Kitt.ai with microphone functionality build-in.☆13Dec 12, 2018Updated 7 years ago
- CogBeacon is a multi-modal dataset designed to target the effects of cognitive fatigue in human performance. The dataset consists of 76 s…☆18Nov 26, 2019Updated 6 years ago
- A neural attention model for speech command recognition☆187Jul 12, 2025Updated last year
- Using spectrograms and convolutional neural networks to listen to environment sounds.☆32Jul 23, 2021Updated 5 years ago
- Classifies percussion audio samples with a CNN-LSTM, written in python and pytorch. Also exports to Drumkv1 (lv2 plugin)☆14Aug 20, 2020Updated 6 years ago
- 🎯 Speech Recognition Challenge by Speech Lab - IIT Madras☆10Nov 5, 2020Updated 5 years ago
- ☆10Apr 2, 2024Updated 2 years ago