A set of scripts that extract speech features (so far MFCCs, FBANKs, STFT, and kinda dominant frequency) and trains CNN, LSTM, or CNN+LSTM models with those features.
☆55Mar 24, 2023Updated 3 years ago
Alternatives and similar repositories for Build-CNN-or-LSTM-or-CNNLSTM-with-speech-features
Users that are interested in Build-CNN-or-LSTM-or-CNNLSTM-with-speech-features are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Calculate MFCC/Fbank feature for wav files☆15Nov 21, 2017Updated 8 years ago
- In this work we propose two postprocessing approaches applying convolutional neural networks (CNNs) either in the time domain or the ceps…☆28Mar 8, 2020Updated 6 years ago
- A repository for emotion recognition from speech, text and mocap data from IEMOCAP dataset☆13Dec 12, 2018Updated 7 years ago
- (tensorflow) Wiener Filter based Speech Enhancement(LSTM/BLSTM, GRU/BGRU, Transformer)☆15Dec 3, 2019Updated 6 years ago
- Speechflow for emotion recognition related information decomposition☆10Jul 27, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Bidirectional LSTM network for speech emotion recognition.☆266Mar 31, 2019Updated 7 years ago
- 用于机器学习的语音特征提取,包含FBank和MFCC等,原理讲解和step by step的实现☆54May 17, 2019Updated 7 years ago
- ☆18Nov 10, 2019Updated 6 years ago
- python codes to extract MFCC and FBANK speech features for Kaldi☆67Nov 28, 2018Updated 7 years ago
- Several studies have been carried out to analyse Parkinson’s disease using speech impairments. Various tools and techniques have been use…☆12Apr 1, 2019Updated 7 years ago
- Deep learning using CNN for Mandarin Chinese tone classification☆39Apr 5, 2019Updated 7 years ago
- In this project, we wish to identify psychiatric disorders through patient's speech☆12Jun 6, 2021Updated 5 years ago
- Blind Source Separation: Independent Component Analysis for EEG data with python-MNE package and SSVEP☆12Nov 26, 2018Updated 7 years ago
- Matlab tools for pathological voice analysis☆14May 12, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆10May 22, 2023Updated 3 years ago
- Inspired by the convolutional recurrent neural network(CRNN) and inception, we propose a multiscale time-frequency convolutional recurren…☆22Apr 15, 2020Updated 6 years ago
- VoiceCode is an Open Source initiative started by the National Research Council of Canada, to develop a programming by voice toolbox. The…☆10Apr 17, 2020Updated 6 years ago
- Masked ConditionaL Neural Networks☆15Jul 6, 2023Updated 3 years ago
- Web app created to collect audios for course project☆10Apr 6, 2018Updated 8 years ago
- Die Webseite des Chaostreff Potsdam☆11May 21, 2026Updated 4 months ago
- simple energy vad☆19Jun 3, 2017Updated 9 years ago
- Ein Programm zur Beschleunigung von Sprachaufnahmen☆12Apr 16, 2018Updated 8 years ago
- [Research] Monaural Speech Enhancement through Wave-U-Net (SEWUNet)☆32Nov 22, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- assignments for e6870 ASR class☆42Apr 23, 2019Updated 7 years ago
- A Python 2.7 implementation of Mel Frequency Cepstral Coefficients (MFCC) and Dynamic Time Warping (DTW) algorithms for Automated Speech …☆16Apr 23, 2018Updated 8 years ago
- 小内存、显存(低于4g)使用bert做下游任务的一个方案☆14Nov 19, 2019Updated 6 years ago
- Front-end speech processing aims at extracting proper features from short- term segments of a speech utterance, known as frames. It is a …☆257Mar 3, 2023Updated 3 years ago
- Supervised Speech Representation Learning for Parkinson's Disease Classification☆18Oct 26, 2021Updated 4 years ago
- Detect Depression with AI Sub-challenge (DSS) of AVEC2019 experienment version via YZK☆15May 28, 2021Updated 5 years ago
- A pipeline from Dataset Gathering,Data annotations, Model training,Model Evaluation for viseme (visual sound phoneme) classification☆16Jan 19, 2021Updated 5 years ago
- Qt for Python workshop☆11Nov 23, 2021Updated 4 years ago
- Acoustic feature extraction using Librosa library and openSMILE toolkit.使用Librosa音频处理库和openSMILE工具包,进行简单的声学特征提取☆217May 26, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Collection space for resources for the @OpenTechSchool data science co-learning meetup☆44Jul 14, 2018Updated 8 years ago
- ICASSP 2021 accepted paper☆19May 20, 2021Updated 5 years ago
- 语音识别 MFCCs特征处理 cnn神经网络☆104Jan 22, 2019Updated 7 years ago
- A set of speech feature extraction functions for ASR and speaker identification written in matlab.☆43Oct 28, 2016Updated 9 years ago
- Lightweight and extensible DTN library☆19Mar 28, 2020Updated 6 years ago
- ☆22Jul 28, 2018Updated 8 years ago
- TensorFlow,DCGAN,VAE,LSTM,CNN,Acoustic Scene Classification☆11Jun 5, 2019Updated 7 years ago