Primer on CTC implementation in pure Python PyTorch code
☆114Jun 22, 2026Updated 2 months ago
Alternatives and similar repositories for ctc
Users that are interested in ctc are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Mar 15, 2022Updated 4 years ago
- Computes the MWER (minimum WER) Loss with CTC beam search. Knowledge distillation for CTC loss.☆60Sep 6, 2023Updated 3 years ago
- ☆16Jun 13, 2022Updated 4 years ago
- This repository contains source codes for SoftCTC. Original paper can be found here: https://arxiv.org/abs/2212.02135☆19Mar 7, 2023Updated 3 years ago
- Segment a given audio into utterances using a trained end-to-end ASR model.☆75Oct 9, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challenge☆16Mar 26, 2022Updated 4 years ago
- A CSRankings-like index for speech researchers☆35Oct 16, 2024Updated last year
- Baseline convolutional ASR system in PyTorch☆21Nov 16, 2023Updated 2 years ago
- A Python interface to OpenFst (fix FstDrawer interface issue for 1.6 version)☆17Apr 2, 2018Updated 8 years ago
- CAT is more than a CRF-based ASR toolkit: it provides a complete workflow for data-efficient end-to-end ASR, supporting CTC, CTC-CRF, RNN…☆369Feb 5, 2026Updated 7 months ago
- Code for DeCoAR (ICASSP 2020) and BERTphone (Odyssey 2020)☆104Nov 26, 2022Updated 3 years ago
- tts fronted-end☆11Dec 19, 2018Updated 7 years ago
- Implementation of Imputer: Sequence Modelling via Imputation and Dynamic Programming in PyTorch☆58May 3, 2020Updated 6 years ago
- Implements of CTC, Speech-Transformer and CIF for end-to-end speech recognition with pytorch☆23Jul 28, 2020Updated 6 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆277Jan 15, 2021Updated 5 years ago
- CUDA-Warp RNN-Transducer☆215Feb 22, 2023Updated 3 years ago
- Levenshtein edit-distance on PyTorch and CUDA☆93Jan 24, 2023Updated 3 years ago
- Applications using the GTN library and code to reproduce experiments in "Differentiable Weighted Finite-State Transducers"☆83Jul 20, 2022Updated 4 years ago
- HMM, CTC, RNN-Transducer, forward-backward algorithm☆20Sep 5, 2023Updated 3 years ago
- ☆28Jan 29, 2021Updated 5 years ago
- Efficient Neural Architecture Search via Straight-Through Gradients☆13Nov 12, 2020Updated 5 years ago
- PyTorch CTC Decoder bindings☆857Apr 4, 2024Updated 2 years ago
- [ASRU 2021] Efficient Conformer: Progressive Downsampling and Grouped Attention for Automatic Speech Recognition☆220Jun 22, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PyTorch implementation of LF-MMI for End-to-end ASR☆221Jan 14, 2021Updated 5 years ago
- Russian phonetical transcription☆11May 20, 2026Updated 3 months ago
- RawNet: Fast End-to-End Neural Vocoder☆42May 29, 2019Updated 7 years ago
- Decoders from Kaldi using OpenFst☆35Apr 10, 2026Updated 4 months ago
- End-to-end ASR/LM implementation with PyTorch☆595Aug 30, 2021Updated 5 years ago
- Segment an audio file and obtain utterance alignments. (Python package)☆348May 15, 2024Updated 2 years ago
- Auto Segmentation Criterion (ASG) implemented in pytorch☆51Oct 1, 2021Updated 4 years ago
- mWER loss implementation in tensorflow☆31Sep 7, 2020Updated 6 years ago
- streaming attention networks for end-to-end automatic speech recognition☆56May 6, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆46Nov 2, 2023Updated 2 years ago
- Syllable Segmentation and Cross-Lingual Generalization in a Visually Grounded, Self-Supervised Speech Model☆35Aug 27, 2023Updated 3 years ago
- a lightweight speech processing toolkit based on Pytorch and (Py)Kaldi☆355Dec 25, 2020Updated 5 years ago
- Losses and decoders for end-to-end ASR and OCR☆34Oct 30, 2020Updated 5 years ago
- Towards hot directions in industrial end to end speech recognition☆329Nov 30, 2021Updated 4 years ago
- Implementation of the AlignTTS☆77Jul 6, 2023Updated 3 years ago
- Connectionist Temporal Classification (CTC) decoding algorithms: best path, beam search, lexicon search, prefix search, and token passing…☆836Jan 31, 2026Updated 7 months ago