Listen, Attend and Spell (LAS) framework for speech recognition (see https://arxiv.org/pdf/1508.01211.pdf).
☆32Jun 27, 2019Updated 7 years ago
Alternatives and similar repositories for LAS-SpeechRecognition
Users that are interested in LAS-SpeechRecognition are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A listen attend and spell reimplementation in tensorflow, using a custom attention mechanism.☆44Apr 24, 2019Updated 7 years ago
- MXNet implementation of RNN Transducer (Graves 2012): Sequence Transduction with Recurrent Neural Networks☆140Jun 7, 2021Updated 5 years ago
- This is the TensorFlow implementation of the Google LAS model.☆14Feb 1, 2019Updated 7 years ago
- ☆43Jun 25, 2018Updated 8 years ago
- End-to-End speech recognition implementation base on TensorFlow (CTC, Attention, and MTL training)☆314Jan 23, 2018Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Tensorflow implementation of "Listen, Attend and Spell" authored by William Chan. This project utilizes input pipeline and estimator API …☆90Jan 31, 2019Updated 7 years ago
- ESPnet extensions for semi-supervised end-to-end speech recognition. See also https://github.com/ShigekiKarita/espnet-semi-supervised/tre…☆38Feb 13, 2020Updated 6 years ago
- simple energy vad☆19Jun 3, 2017Updated 9 years ago
- Code for end-to-end ASR with neural networks, build with TensorFlow☆110Jan 24, 2019Updated 7 years ago
- Small-footprint Keyword Spotting☆18Jul 28, 2019Updated 7 years ago
- Automatic Speech Recognition with TensorFlow(CNN+BLSTM+CTC)☆12Aug 9, 2018Updated 8 years ago
- This is an open source project (formerly named Listen, Attend and Spell - PyTorch Implementation) for end-to-end ASR implemented with Pyt…☆1,207Dec 19, 2020Updated 5 years ago
- audio cfeatures extraction tool from wav to h5features format☆19May 24, 2019Updated 7 years ago
- ☆13Sep 12, 2017Updated 9 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- End-to-End Attention-Based Large Vocabulary Speech Recognition☆265Nov 22, 2022Updated 3 years ago
- Speech command classification on Speech-Command v0.02 dataset using PyTorch and torchaudio. In this example, three models have been train…☆10Dec 5, 2022Updated 3 years ago
- Implementation of a speaker identification and a speaker verification system based on Gaussian Mixture Models (GMM) in combination with a…☆21Mar 1, 2018Updated 8 years ago
- Reproduction of a paper"Small-footprint keyword spotting using deep neural networks"☆12Mar 11, 2019Updated 7 years ago
- Conv-LSTM-CTC speech recognition network (end-to-end), written in TensorFlow.☆72Mar 21, 2019Updated 7 years ago
- solutions for https://www.kaggle.com/c/tensorflow-speech-recognition-challenge☆31Jan 28, 2018Updated 8 years ago
- ☆12Apr 1, 2020Updated 6 years ago
- Listen, Attend and Spell - PyTorch Implementation☆17Dec 28, 2018Updated 7 years ago
- The implementation of 'Watch, Listen, Attend and Spell’ (WLAS) network that learns to transcribe videos of mouth motion to character on p…☆11Mar 23, 2018Updated 8 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Just that (for Matlab)☆18May 18, 2012Updated 14 years ago
- OpenPose: A Real-Time Multi-Person Keypoint Detection And Multi-Threading C++ Library☆11Jul 13, 2017Updated 9 years ago
- ☆30Nov 9, 2018Updated 7 years ago
- Triangle Attack: A Query-efficient Decision-based Adversarial Attack (ECCV 2022)☆16Jul 19, 2022Updated 4 years ago
- ☆55Jun 15, 2020Updated 6 years ago
- THEANO-KALDI-RNNs is a project implementing various Recurrent Neural Networks (RNNs) for RNN-HMM speech recognition. The Theano Code is c…☆35Apr 15, 2018Updated 8 years ago
- An implementation of rnn transducer for sequence labeling problem☆22Feb 24, 2018Updated 8 years ago
- Use speech_to_text for keyword search in audio files.☆12May 5, 2021Updated 5 years ago
- Homemade LightGBM and VGG-net experiment setup for DCASE2017 task 1☆11Aug 8, 2017Updated 9 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Python implementation of pre-processing for End-to-End speech recognition☆70Feb 19, 2018Updated 8 years ago
- Implements of CTC, Speech-Transformer and CIF for end-to-end speech recognition with pytorch☆23Jul 28, 2020Updated 6 years ago
- https://www.kaggle.com/c/tensorflow-speech-recognition-challenge/☆21Mar 1, 2018Updated 8 years ago
- The RWTH extensible training framework for universal recurrent neural networks☆377Updated this week
- [COLM '25] Single-Pass Document Scanning for Question Answering☆15Sep 6, 2026Updated 2 weeks ago
- PyTorch implementation of Listen Attend and Spell Automatic Speech Recognition (ASR).☆39Jul 25, 2019Updated 7 years ago
- MobileNet trained with VoxCeleb dataset and used for voice verification☆18Oct 26, 2022Updated 3 years ago