speech-to-text in pytorch
☆83Mar 14, 2019Updated 7 years ago
Alternatives and similar repositories for patter
Users that are interested in patter are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆76Mar 18, 2022Updated 4 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- A PyTorch implementation of speech recognition based on DeepMind's WaveNet☆18Jun 5, 2018Updated 8 years ago
- Yet another speech toolkit based on Kaldi and PyTorch☆173Jul 1, 2020Updated 6 years ago
- PyTorch implementation of LF-MMI for End-to-end ASR☆221Jan 14, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A fully convolution-network for speech-to-text, built on pytorch.☆126May 20, 2020Updated 6 years ago
- A KALDI/C++ implementation of GoogleBrain's SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition☆15Sep 4, 2019Updated 6 years ago
- Data and code related to the ICASSP submission "A comparison of methods for OOV-word recognition"☆17Nov 28, 2021Updated 4 years ago
- Speech Recognition using DeepSpeech2.☆17Nov 19, 2019Updated 6 years ago
- Kaldi model converter to ONNX☆248Jan 27, 2023Updated 3 years ago
- The magic message bus☆15Mar 21, 2019Updated 7 years ago
- An example directory for running Multi-Task Learning training on Kaldi neural networks. In Kaldi-speak, this is an egs dir for nnet3 trai…☆55Jan 2, 2020Updated 6 years ago
- SpeechNAS-Better-Trade-off-between-Latency-and-Accuracy-for-Large-Scale-Speaker-Verification☆30Mar 24, 2023Updated 3 years ago
- ☆22Sep 24, 2018Updated 7 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- An implementation of the Wav2Letter Speech-to-Text model using PyTorch.☆14Mar 8, 2023Updated 3 years ago
- Code for end-to-end ASR with neural networks, build with TensorFlow☆110Jan 24, 2019Updated 7 years ago
- An efficient implementation of RNN-T Prefix Beam Search in C++/CUDA.☆67Jan 7, 2026Updated 7 months ago
- Interspeech 2019 tutorial materials☆49Sep 26, 2019Updated 6 years ago
- HMM, CTC, RNN-Transducer, forward-backward algorithm☆20Sep 5, 2023Updated 2 years ago
- A Chaos Monkey-like failure recovery testing tool for applications running on Marathon☆15Jul 22, 2015Updated 11 years ago
- CNN and LSTM multi-label text classification☆11Apr 13, 2018Updated 8 years ago
- AMR Eager trained on English, German, Italian, Spanish and Chinese☆24May 29, 2020Updated 6 years ago
- Custom decoders for Kaldi☆13Jun 5, 2019Updated 7 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Program to benchmark various speech recognition APIs☆82Sep 6, 2019Updated 6 years ago
- Custom decoders for Kaldi☆81Jun 10, 2019Updated 7 years ago
- CTC+Beam_Search+kenlm 是用于以汉字为声学模型建模单元的解码系统☆49Jun 27, 2018Updated 8 years ago
- ☆14Jun 12, 2015Updated 11 years ago
- INTERSPEECH 2019 Tutorial Materials☆194Mar 30, 2021Updated 5 years ago
- End-to-End Automatic Speech Recognition on PyTorch☆304Jun 2, 2022Updated 4 years ago
- Keyword Search Recipe for Subword ASR☆30Jul 12, 2019Updated 7 years ago
- Deep learning based speech source separation using Pytorch☆319Nov 20, 2020Updated 5 years ago
- PyTorch implementation of TinyWASE described in our paper "Compressing Speaker Extraction Model with Ultra-low Precision Quantization and…☆11Jun 28, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- readers that enable reading kaldi ark in tensorflow☆17Mar 7, 2018Updated 8 years ago
- Recurrent Neural Aligner☆50Apr 14, 2020Updated 6 years ago
- CAT is more than a CRF-based ASR toolkit: it provides a complete workflow for data-efficient end-to-end ASR, supporting CTC, CTC-CRF, RNN…☆369Feb 5, 2026Updated 6 months ago
- A list of publically available audio data that anyone can download for ASR or other speech activities☆237Aug 6, 2021Updated 5 years ago
- Sound augmentation using Large-scale audio dataset (Audioset)☆45Jun 29, 2021Updated 5 years ago
- STT Service based on Kaldi ASR☆15Aug 17, 2018Updated 8 years ago
- pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch,…☆2,404Mar 14, 2022Updated 4 years ago