Classification of 11 types of audio clips using MFCCs features and LSTM. Pretrained on Speech Command Dataset with intensive data augmentation.
☆43Dec 14, 2022Updated 3 years ago
Alternatives and similar repositories for Speech-Commands-Classification-by-LSTM-PyTorch
Users that are interested in Speech-Commands-Classification-by-LSTM-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Speech command classification on Speech-Command v0.02 dataset using PyTorch and torchaudio. In this example, three models have been train…☆10Dec 5, 2022Updated 3 years ago
- In this repository, I implement a system for detecting specific spoken words in speech signal. When reading a speech signal, I detect not…☆19Sep 27, 2021Updated 4 years ago
- ☆21Mar 8, 2020Updated 6 years ago
- Speech Commands Recognition using end-to-end deep learning models in pytorch☆28Oct 8, 2020Updated 5 years ago
- Speech command recognition with capsule network & various NNs / KWS on Google Speech Command Dataset.☆25Jan 28, 2019Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A CNN audio classifier via spectrogram images.☆10Jul 21, 2017Updated 9 years ago
- ☆10May 22, 2023Updated 3 years ago
- VoiceCode is an Open Source initiative started by the National Research Council of Canada, to develop a programming by voice toolbox. The…☆10Apr 17, 2020Updated 6 years ago
- Triangle Attack: A Query-efficient Decision-based Adversarial Attack (ECCV 2022)☆16Jul 19, 2022Updated 4 years ago
- Repository for code and paper submitted for APSIPA 2019, Lanzhou, China☆21Aug 2, 2024Updated last year
- Speech commands recognition with PyTorch | Kaggle 10th place solution in TensorFlow Speech Recognition Challenge☆201Jan 19, 2024Updated 2 years ago
- 3D reconstruction of a scene from a sequence of RGB and depth images in Matlab☆14Feb 16, 2020Updated 6 years ago
- Speech Recognition for speakers with speech disorders due to diseases like Cerebral Palsy, Parkinson or Amyotrophic Lateral Sclerosis ALS…☆23Mar 26, 2017Updated 9 years ago
- Repository for Weak Label Learning for Audio Events - A closer look. Uses Audioset subset data provided for reproducibility.☆32Sep 13, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A Python 2.7 implementation of Mel Frequency Cepstral Coefficients (MFCC) and Dynamic Time Warping (DTW) algorithms for Automated Speech …☆16Apr 23, 2018Updated 8 years ago
- Classify the emotions from variable-length speech segments☆11Mar 29, 2018Updated 8 years ago
- Implementation of a standard synthetic speech detection approach via bob.bio.spear Python package☆16Mar 23, 2017Updated 9 years ago
- NightSurveillance Sataset for Pedestrian Detection☆11Jul 30, 2020Updated 5 years ago
- Signal processing and clustering of EEG signals☆11Sep 13, 2018Updated 7 years ago
- https://dodiku.github.io/audio_noise_clustering/results/ ==> An experiment with a variety of clustering (and clustering-like) techniques …☆26May 5, 2017Updated 9 years ago
- System for Emotion Detection in given speech data using joint modelling of hand crafted prosody rich features , MFCC features and LSTM ba…☆10Nov 15, 2017Updated 8 years ago
- Speech Recognition in python☆10Jul 12, 2018Updated 8 years ago
- Mispronunciation detection code for jingju singing voice☆19Sep 5, 2018Updated 7 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆22Jul 28, 2018Updated 8 years ago
- Attention-based model for keywords spotting☆19Aug 9, 2021Updated 4 years ago
- Keyword spotting for audio with attention (KWS model for audio)☆18Jul 15, 2021Updated 5 years ago
- Music IR Library for Python☆13Nov 18, 2015Updated 10 years ago
- numpy implementation of Recurrent Neural Network☆41Jun 2, 2019Updated 7 years ago
- This repository is developed in MATLAB. Speech Augmentation is based on Adaptive Filtering while Endpoint Detection is based on Voice Act…☆10Dec 7, 2020Updated 5 years ago
- Predicting Billboard's Year-End Hot 100 Songs using audio features from Spotify and lyrics from Musixmatch☆18Jul 14, 2024Updated 2 years ago
- A set of speech feature extraction functions for ASR and speaker identification written in matlab.☆43Oct 28, 2016Updated 9 years ago
- Classify various radiology images into respective categories.Being done using various shape and texture features for feature extraction a…☆10May 20, 2016Updated 10 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repository provides a small Python wrapper for the Matlab tool SNR Eval provided by Labrosa: https://labrosa.ee.columbia.edu/project…☆12Jun 22, 2022Updated 4 years ago
- Recognizing common speech commands using Keras and Tensorflow.☆10Dec 17, 2018Updated 7 years ago
- Some PyTorch code for the Kaggle Speech Recognition Challenge☆13Feb 7, 2019Updated 7 years ago
- Flask webapp/endpoint that compares the user's speech with different accents and assigns similarity scores based on speed, voice (DTW/MFC…☆18Jun 27, 2017Updated 9 years ago
- 可以将bert相关模型在网站上进行展示、测试,后端使用flask,前端使用vue。☆11Aug 20, 2021Updated 4 years ago
- my approach to the kaggle speech recognition challenge☆25May 11, 2018Updated 8 years ago
- Acoustic and language models for minorised languages.☆26Jul 17, 2026Updated last week