This repository reports how to build a speech to text model to recognize short commands. Best of all, developing and including speech recognition in a Python project using Keras is really simple.
☆23Jul 21, 2020Updated 6 years ago
Alternatives and similar repositories for speech2text_keras
Users that are interested in speech2text_keras are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository contains code for a tutorial on end to end automatic speech recognition.☆18Sep 10, 2019Updated 6 years ago
- A Generative Adversarial Network for Shakuhachi Music☆14Jul 2, 2019Updated 7 years ago
- Recognizing common speech commands using Keras and Tensorflow.☆10Dec 17, 2018Updated 7 years ago
- implement end-to-end asr algorithm with tensorflow☆40Aug 23, 2018Updated 8 years ago
- To calculate the BLUE score☆11Jun 7, 2016Updated 10 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Speech recognition system implemented using tensorflow☆16Feb 2, 2023Updated 3 years ago
- Speech recognition with CTC in Keras with Tensorflow backend☆31Mar 24, 2023Updated 3 years ago
- Speech recognition framework using keras☆13May 18, 2018Updated 8 years ago
- This repository contains all the code necessary for running the multilingual distilwhisper from Ferraz et al. 2024 IEEE ICASSP paper.☆34Apr 22, 2026Updated 4 months ago
- Repository for my paper: Dimensional Speech Emotion Recognition Using Acoustic Features and Word Embeddings using Multitask Learning☆17Aug 2, 2024Updated 2 years ago
- Code for the paper: Deep Residual Networks with Auditory Inspired Features for Robust Speech Recognition.☆21Mar 22, 2017Updated 9 years ago
- Baidu's CTC Decoders, including Greedy, Beam Search and Beam Search with KenLM Language Model☆24Oct 28, 2023Updated 2 years ago
- ☆11Jun 4, 2021Updated 5 years ago
- A Simple Automatic Speech Recognition (ASR) Model in Tensorflow, which only needs to focus on Deep Neural Network. It's easy to test popu…☆20Jan 18, 2018Updated 8 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Shows how to encrypt data held in public space☆11Aug 11, 2017Updated 9 years ago
- Comparing Audio Features for Unsupervised Sound Classification☆10Jun 22, 2022Updated 4 years ago
- Speaker recognition system based upon classification of Mel-Frequency Cepstral Coefficients (MFCC) using a minimum-distance classifier and…☆20Sep 15, 2010Updated 15 years ago
- 🗣️ Convert between phonetic alphabets☆11Feb 7, 2022Updated 4 years ago
- Scripts to convert audio files to spectrograms and back☆12Nov 23, 2017Updated 8 years ago
- WaveGANによる音声生成器☆13Feb 9, 2024Updated 2 years ago
- C/C++ preprocessor-like tool for a range of languages (i.e., #ifdef, #ifndef, #if-else, #include, etc. for Python, LaTeX, Bash, JavaScrip…☆11Jan 6, 2021Updated 5 years ago
- CNN based Minimal model for recognizing word☆61May 7, 2018Updated 8 years ago
- ☆13Aug 25, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Multi-lingual AudioCaps☆14Nov 20, 2023Updated 2 years ago
- Generates Google Drive shareable links for PDFs and links them to a group Zotero database☆13Oct 25, 2020Updated 5 years ago
- Development of an R package for metacognition researchers☆13Jun 2, 2026Updated 2 months ago
- Implementaion RNN tranceducer☆23Jun 25, 2019Updated 7 years ago
- Rainbowgram with Python☆13Jan 28, 2019Updated 7 years ago
- ☆19May 9, 2019Updated 7 years ago
- ☆11Nov 29, 2020Updated 5 years ago
- Dimensionality reduction (UMAP, t-SNE, PCA) for ImageJ/Fiji☆12May 6, 2025Updated last year
- Keras implementation of conditional waveGAN. Application to knocking sound effects with emotion.