Sequence Modelling with CTC
☆52Dec 29, 2022Updated 3 years ago
Alternatives and similar repositories for post--ctc
Users that are interested in post--ctc are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Recurrent Neural Aligner☆51Apr 14, 2020Updated 6 years ago
- Corpus of oral arguments (recorded speech + official transcripts) of the United States Supreme Court☆22Dec 8, 2022Updated 3 years ago
- Self-Supervised Speech Pre-training and Representation Learning Toolkit.☆10Feb 29, 2024Updated 2 years ago
- compare three CTC decoder, that is greedy decoder, beam decoder and prefix beam decoder☆20Jul 10, 2018Updated 8 years ago
- A small C++ library for efficient calculation of rotation invariant features in 2D images using OpenCV.☆12Feb 12, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- a catch-all repo☆11Dec 28, 2023Updated 2 years ago
- Pybind11 bindings for Kaldi☆15Jul 11, 2026Updated 2 weeks ago
- Transfer learning approach to pronunciation scoring☆12Jan 17, 2024Updated 2 years ago
- Proposed splits for the LREC Wikipron paper☆15Apr 7, 2020Updated 6 years ago
- A Weakly Supervised Forced Alignment for disluent speech☆15Nov 12, 2023Updated 2 years ago
- Phonetically-Oriented Word Error Rate☆36May 4, 2019Updated 7 years ago
- PyTorch implementation of LF-MMI for End-to-end ASR☆221Jan 14, 2021Updated 5 years ago
- Command line tool for forced-alignment of Spanish speech data☆13Dec 31, 2025Updated 6 months ago
- This repo is for residual-connected sentence encoder for NLI.☆11Jan 21, 2018Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Mason-Alberta Phonetic Segmenter☆15Feb 24, 2026Updated 5 months ago
- Standard libraries for audio processing, especially STFT and Spherical Harmonics decomposition of a soundfield.☆10Nov 29, 2021Updated 4 years ago
- speech recognition using Kaldi framework☆12Dec 25, 2019Updated 6 years ago
- A fast parallel implementation of RNN Transducer.☆12Apr 8, 2025Updated last year
- pronunciation LEXicons for Any Low-resource Language☆21Jul 14, 2020Updated 6 years ago
- Code for the Paper 'On the Connection Between Adversarial Robustness and Saliency Map Interpretability' by C. Etmann, S. Lunz, P. Maass, …☆16May 9, 2019Updated 7 years ago
- Library about construction helper for Generative models e.g. Flow-based Model with Tensorflow 2.x.☆12Feb 16, 2023Updated 3 years ago
- 该脚本根据语料文件生成对应的图像文件,适用于文本识别等CV任务☆29Aug 4, 2021Updated 4 years ago
- [WIP] Unofficial Implementation of Microsoft's PromptTTS2☆53Oct 31, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆13Dec 4, 2017Updated 8 years ago
- ☆19Jun 28, 2022Updated 4 years ago
- A simple shoot 'em ups style game using Rust's Bevy crate https://play.marcoinacio.com☆15Feb 25, 2024Updated 2 years ago
- Flutter Bridge for .NET Maui☆13Jul 12, 2024Updated 2 years ago
- Easy to install cross-platform python desktop app that gets video via OpenCV and displays it via LGPL Qt 5 for Python (PySide2) GUI compo…☆10Jul 18, 2019Updated 7 years ago
- A toy-like Text-to-Speech for Chinese/Mandarin synthesize, inspired by Tacotron & FastSpeech2 & RefineGAN.☆15May 25, 2022Updated 4 years ago
- ☆13May 9, 2022Updated 4 years ago
- Koel Labs innovates open-source speech research, inclusive speech technologies, and real-time pronunciation feedback for language learner…☆24Jul 13, 2026Updated last week
- A particle swarm optimization library created by Numenta for hyperparameter optimization.☆18Aug 18, 2015Updated 10 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SinGlow is a part of my Singing voice synthesis system. It can extract features of sound, particularly songs and musics. Then we can use …☆11Oct 9, 2021Updated 4 years ago
- Gallery of applications built using bqplot and widget libraries like ipywidgets, ipydatagrid etc.☆11Feb 1, 2023Updated 3 years ago
- Iterative Closest Point algorithm for scans/mesh alignment (with subsampling and point to plane improvements).☆10Jul 15, 2018Updated 8 years ago
- End-to-end speech recognition using TensorFlow☆48Apr 2, 2018Updated 8 years ago
- Connectionist Temporal Classification (CTC) decoding algorithms: best path, beam search, lexicon search, prefix search, and token passing…☆837Jan 31, 2026Updated 5 months ago
- DatasetImgLabeler is a image annotation tool for researchers to prepare datasets in ICDAR2015 format☆12Dec 7, 2019Updated 6 years ago
- Visual Search in Natural Scenes benchmark☆20Sep 19, 2024Updated last year