Classification of 11 types of audio clips using MFCCs features and LSTM. Pretrained on Speech Command Dataset with intensive data augmentation.
☆43Dec 14, 2022Updated 3 years ago
Alternatives and similar repositories for Speech-Commands-Classification-by-LSTM-PyTorch
Users that are interested in Speech-Commands-Classification-by-LSTM-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Speech command classification on Speech-Command v0.02 dataset using PyTorch and torchaudio. In this example, three models have been train…☆10Dec 5, 2022Updated 3 years ago
- In this repository, I implement a system for detecting specific spoken words in speech signal. When reading a speech signal, I detect not…☆19Sep 27, 2021Updated 4 years ago
- ☆21Mar 8, 2020Updated 6 years ago
- Speech Commands Recognition using end-to-end deep learning models in pytorch☆28Oct 8, 2020Updated 5 years ago
- Keyword Spotting for detecting a word in an audio file☆17Jul 21, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- UrbanSound8K dataset classification using MLP and CNN☆11Nov 17, 2021Updated 4 years ago
- A CNN audio classifier via spectrogram images.☆10Jul 21, 2017Updated 9 years ago
- ☆10May 22, 2023Updated 3 years ago
- VoiceCode is an Open Source initiative started by the National Research Council of Canada, to develop a programming by voice toolbox. The…☆10Apr 17, 2020Updated 6 years ago
- Repository for code and paper submitted for APSIPA 2019, Lanzhou, China☆21Aug 2, 2024Updated 2 years ago
- [ACL'22] Training-free Neural Architecture Search for RNNs and Transformers☆14May 26, 2024Updated 2 years ago
- Speech commands recognition with PyTorch | Kaggle 10th place solution in TensorFlow Speech Recognition Challenge☆201Jan 19, 2024Updated 2 years ago
- 3D reconstruction of a scene from a sequence of RGB and depth images in Matlab☆14Feb 16, 2020Updated 6 years ago
- Repository for Weak Label Learning for Audio Events - A closer look. Uses Audioset subset data provided for reproducibility.☆32Sep 13, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Classify the emotions from variable-length speech segments☆11Mar 29, 2018Updated 8 years ago
- Implementation of a standard synthetic speech detection approach via bob.bio.spear Python package☆16Mar 23, 2017Updated 9 years ago
- Google Speech Command Dataset Classification Neural Network, CNN, RNN☆26Aug 29, 2017Updated 9 years ago
- https://dodiku.github.io/audio_noise_clustering/results/ ==> An experiment with a variety of clustering (and clustering-like) techniques …☆26May 5, 2017Updated 9 years ago
- ☆16Mar 9, 2021Updated 5 years ago
- ☆22Jul 28, 2018Updated 8 years ago
- System for Emotion Detection in given speech data using joint modelling of hand crafted prosody rich features , MFCC features and LSTM ba…☆10Nov 15, 2017Updated 8 years ago
- Mispronunciation detection code for jingju singing voice☆19Sep 5, 2018Updated 8 years ago
- Attention-based model for keywords spotting☆19Aug 9, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Keyword spotting for audio with attention (KWS model for audio)☆18Jul 15, 2021Updated 5 years ago
- Music IR Library for Python☆13Nov 18, 2015Updated 10 years ago
- numpy implementation of Recurrent Neural Network☆41Jun 2, 2019Updated 7 years ago
- This repository is developed in MATLAB. Speech Augmentation is based on Adaptive Filtering while Endpoint Detection is based on Voice Act…☆10Dec 7, 2020Updated 5 years ago
- Predicting Billboard's Year-End Hot 100 Songs using audio features from Spotify and lyrics from Musixmatch☆18Jul 14, 2024Updated 2 years ago
- A set of speech feature extraction functions for ASR and speaker identification written in matlab.☆43Oct 28, 2016Updated 9 years ago
- Classify various radiology images into respective categories.Being done using various shape and texture features for feature extraction a…☆10May 20, 2016Updated 10 years ago
- Recognizing common speech commands using Keras and Tensorflow.☆10Dec 17, 2018Updated 7 years ago
- Some PyTorch code for the Kaggle Speech Recognition Challenge☆13Feb 7, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Flask webapp/endpoint that compares the user's speech with different accents and assigns similarity scores based on speed, voice (DTW/MFC…☆18Jun 27, 2017Updated 9 years ago
- my approach to the kaggle speech recognition challenge☆25May 11, 2018Updated 8 years ago
- CANVAS: Commonsense-Aware Navigation System for Intuitive Human-Robot Interaction [ICRA 2025]☆18Oct 20, 2025Updated 10 months ago
- 2D/3D physics engine for games written in Rust☆12Mar 7, 2022Updated 4 years ago
- CogBeacon is a multi-modal dataset designed to target the effects of cognitive fatigue in human performance. The dataset consists of 76 s…☆18Nov 26, 2019Updated 6 years ago
- A neural attention model for speech command recognition☆187Jul 12, 2025Updated last year
- Subin An's CV☆12Nov 12, 2022Updated 3 years ago