End-to-end speech recognition using TensorFlow
☆48Apr 2, 2018Updated 8 years ago
Alternatives and similar repositories for deepSpeech2
Users that are interested in deepSpeech2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A keras layer implementation of Peddinti's paper "A time delay neural network architecture for efficient modeling of long temporal conte…☆13Nov 19, 2018Updated 7 years ago
- Long audio alignment using Kaldi☆23Apr 22, 2021Updated 5 years ago
- A collection of useful tools for handling speech recognition data☆30Nov 28, 2022Updated 3 years ago
- Recurrent neural network for audio noise reduction☆12Aug 18, 2022Updated 4 years ago
- mixlingual speech recognition system; hybrid (GMM+NNet) model; Kaldi + Keras☆71Nov 20, 2017Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Speech Recognition using DeepSpeech2.☆2,136Dec 13, 2022Updated 3 years ago
- This is now the official location of the Kaldi project.☆24Nov 13, 2019Updated 6 years ago
- RAG code sorting search, RAG knowledge organization☆16Nov 22, 2024Updated last year
- A module for normalising text.☆10Nov 6, 2019Updated 6 years ago
- Some notes on Kaldi☆32Feb 20, 2015Updated 11 years ago
- the repaired code of paper "Age Progression/Regression by Conditional Adversarial Autoencoder---CVPR 2017"☆10Sep 19, 2017Updated 9 years ago
- the infinite ramble in rust, powered by tensorflow. (mfcc cosine similarity matching)☆13Apr 30, 2018Updated 8 years ago
- Decoders from Kaldi using OpenFst☆35Apr 10, 2026Updated 5 months ago
- https://www.kaggle.com/c/siim-acr-pneumothorax-segmentation☆11Sep 11, 2019Updated 7 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The History of Speech Recognition to the Year 2030☆13Aug 14, 2021Updated 5 years ago
- For further understanding the wide array of emotions embedded in human speech, we are introducing an emotional speech corpus. In contrast…☆11Oct 29, 2018Updated 7 years ago
- Deep Neural Network Compression based on Student-Teacher Network☆14Jul 6, 2023Updated 3 years ago
- Tool for creation, manipulation and maintenance of voice corpora☆82May 3, 2024Updated 2 years ago
- Baidu's DeepSpeech updated for better training☆23Sep 5, 2018Updated 8 years ago
- Speech command recognition DenseNet transfer learning from UrbanSound8k in keras tensorflow☆17Jan 19, 2018Updated 8 years ago
- Train a Deep Learning model to classify audio embeddings on IBM's Deep Learning as a Service (DLaaS) platform - Watson Machine Learning☆102Sep 17, 2025Updated last year
- DeepSpeech neon implementation☆220Jan 3, 2023Updated 3 years ago
- python wrapper for opencv gpu pyramid-lk-flow and tv-l1-flow☆10Jul 15, 2017Updated 9 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Sequence Modelling with CTC☆52Dec 29, 2022Updated 3 years ago
- ipython notebooks for feature extraction and training of audio event classifier on ESC-50 dataset.☆10Mar 16, 2018Updated 8 years ago
- SLT 2024 Mandarin Stuttering Event Detection and Automatic Speech Recognition Challenge☆12Jun 11, 2024Updated 2 years ago
- Some simple wrappers around kaldi-asr intended to make using kaldi's (online) decoders as convenient as possible.☆169Feb 23, 2021Updated 5 years ago
- My 1st place solution to the Kaggle Invasive Species Monitoring Competition☆10Aug 17, 2017Updated 9 years ago
- Script to train a German n-gram Language Model on articles of Wikipedia☆14Oct 20, 2018Updated 7 years ago
- voice changer with DNN & GAN on Keras☆13Mar 9, 2018Updated 8 years ago
- MelGAN and Tacotron 2 in PyTorch☆11Oct 22, 2019Updated 6 years ago
- DeepSpeech, Speech To Text, ASR, Speech recognition, Keras, Tensorflow☆30Jan 16, 2018Updated 8 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆13Oct 14, 2020Updated 5 years ago
- A living document for all things Common Voice.☆14Jun 24, 2024Updated 2 years ago
- ☆28Apr 24, 2026Updated 4 months ago
- Implementation of the paper "BERTphone: Phonetically-aware Encoder Representations for Utterance-level Speaker and Language Recognition"☆17Dec 10, 2020Updated 5 years ago
- Code for https://arxiv.org/abs/1712.00254☆18Dec 6, 2017Updated 8 years ago
- Speech command recognition with capsule network & various NNs / KWS on Google Speech Command Dataset.☆25Jan 28, 2019Updated 7 years ago
- CROWN: A Neural Network Verification Framework for Networks with General Activation Functions☆39Dec 13, 2018Updated 7 years ago