speech-to-text in pytorch
☆83Mar 14, 2019Updated 7 years ago
Alternatives and similar repositories for patter
Users that are interested in patter are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆76Mar 18, 2022Updated 4 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- Yet another speech toolkit based on Kaldi and PyTorch☆173Jul 1, 2020Updated 6 years ago
- PyTorch implementation of LF-MMI for End-to-end ASR☆221Jan 14, 2021Updated 5 years ago
- A fully convolution-network for speech-to-text, built on pytorch.☆126May 20, 2020Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A KALDI/C++ implementation of GoogleBrain's SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition☆15Sep 4, 2019Updated 6 years ago
- Data and code related to the ICASSP submission "A comparison of methods for OOV-word recognition"☆17Nov 28, 2021Updated 4 years ago
- Speech Recognition using DeepSpeech2.☆17Nov 19, 2019Updated 6 years ago
- Kaldi model converter to ONNX☆248Jan 27, 2023Updated 3 years ago
- The magic message bus☆15Mar 21, 2019Updated 7 years ago
- An example directory for running Multi-Task Learning training on Kaldi neural networks. In Kaldi-speak, this is an egs dir for nnet3 trai…☆55Jan 2, 2020Updated 6 years ago
- SpeechNAS-Better-Trade-off-between-Latency-and-Accuracy-for-Large-Scale-Speaker-Verification☆30Mar 24, 2023Updated 3 years ago
- ☆22Sep 24, 2018Updated 7 years ago
- An implementation of the Wav2Letter Speech-to-Text model using PyTorch.☆14Mar 8, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Code for end-to-end ASR with neural networks, build with TensorFlow☆110Jan 24, 2019Updated 7 years ago
- An efficient implementation of RNN-T Prefix Beam Search in C++/CUDA.☆67Jan 7, 2026Updated 6 months ago
- Interspeech 2019 tutorial materials☆49Sep 26, 2019Updated 6 years ago
- HMM, CTC, RNN-Transducer, forward-backward algorithm☆20Sep 5, 2023Updated 2 years ago
- A Chaos Monkey-like failure recovery testing tool for applications running on Marathon☆15Jul 22, 2015Updated 11 years ago
- CNN and LSTM multi-label text classification☆11Apr 13, 2018Updated 8 years ago
- Custom decoders for Kaldi☆13Jun 5, 2019Updated 7 years ago
- Program to benchmark various speech recognition APIs☆82Sep 6, 2019Updated 6 years ago
- Custom decoders for Kaldi☆81Jun 10, 2019Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- CTC+Beam_Search+kenlm 是用于以汉字为声学模型建模单元的解码系统☆49Jun 27, 2018Updated 8 years ago
- ☆14Jun 12, 2015Updated 11 years ago
- INTERSPEECH 2019 Tutorial Materials☆194Mar 30, 2021Updated 5 years ago
- Keyword Search Recipe for Subword ASR☆30Jul 12, 2019Updated 7 years ago
- End-to-End Automatic Speech Recognition on PyTorch☆304Jun 2, 2022Updated 4 years ago
- Deep learning based speech source separation using Pytorch☆319Nov 20, 2020Updated 5 years ago
- PyTorch implementation of TinyWASE described in our paper "Compressing Speaker Extraction Model with Ultra-low Precision Quantization and…☆11Jun 28, 2021Updated 5 years ago
- readers that enable reading kaldi ark in tensorflow☆17Mar 7, 2018Updated 8 years ago
- CAT is more than a CRF-based ASR toolkit: it provides a complete workflow for data-efficient end-to-end ASR, supporting CTC, CTC-CRF, RNN…☆368Feb 5, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Recurrent Neural Aligner☆51Apr 14, 2020Updated 6 years ago
- A list of publically available audio data that anyone can download for ASR or other speech activities☆237Aug 6, 2021Updated 4 years ago
- Sound augmentation using Large-scale audio dataset (Audioset)☆45Jun 29, 2021Updated 5 years ago
- STT Service based on Kaldi ASR☆15Aug 17, 2018Updated 7 years ago
- pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch,…☆2,401Mar 14, 2022Updated 4 years ago
- Espresso: A Fast End-to-End Neural Speech Recognition Toolkit☆939Sep 4, 2024Updated last year
- mixlingual speech recognition system; hybrid (GMM+NNet) model; Kaldi + Keras☆71Nov 20, 2017Updated 8 years ago