An automatic speech recognition API
☆84Jun 26, 2026Updated 3 weeks ago
Alternatives and similar repositories for linto-stt
Users that are interested in linto-stt are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Build a LinTO OS Image which boots on Raspberry Pi3☆14Jul 8, 2020Updated 6 years ago
- Transcription and annotation interface for recorded audio or video files☆58Updated this week
- Speaker diarization service☆27Jul 2, 2026Updated 2 weeks ago
- GUI Tool to create, manage and test Keyword Spotting models using TF 2.0☆13Feb 1, 2021Updated 5 years ago
- Continual pretraining of foundation LLM using ⚡ Lightning Fabric☆37Nov 27, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Buildings block for voice-enabled applications in the browser☆37Feb 5, 2026Updated 5 months ago
- This repository provides data and code for "Vox Populi, Vox DIY: Benchmark Dataset for Crowdsourced Audio Transcription" paper.☆16Jul 22, 2021Updated 4 years ago
- Code for continual pretraining of LUCIE☆52Jun 2, 2026Updated last month
- AMI and ICSI Corpora in JSON format.☆38Sep 29, 2023Updated 2 years ago
- ☆13Aug 7, 2021Updated 4 years ago
- phone inventory library☆17May 15, 2023Updated 3 years ago
- enhan(t) is an open source toolkit which enables you to enhance the web experience of existing video conferencing solutions like Zoom, MS…☆15Apr 28, 2022Updated 4 years ago
- Thai Grapheme to Phoneme (G2P) Wiktionary Corpus☆13Jul 25, 2022Updated 3 years ago
- TTS for Singlish using Tacotron2, the IMDA corpus, and Pachyderm.☆11Jan 11, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Next-generation Punkt sentence boundary detection with zero dependencies☆32Nov 18, 2025Updated 8 months ago
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.☆13Feb 13, 2021Updated 5 years ago
- ☆17Apr 14, 2023Updated 3 years ago
- ☆11Sep 5, 2025Updated 10 months ago
- Model for recasing and repunctuating ASR transcripts☆141Apr 10, 2024Updated 2 years ago
- ☆15Updated this week
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆34Sep 25, 2025Updated 9 months ago
- radiomixer☆14Feb 16, 2022Updated 4 years ago
- ☆17Jun 16, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Character-level conversion between Hebrew text and Latin transliteration using deep learning - a demonstration of seq2seq training.☆16Jun 27, 2023Updated 3 years ago
- DEPRECATED - A webapp for collecting speech samples for voice recognition testing and training☆20May 23, 2019Updated 7 years ago
- Arabic Grapheme-to-Phoneme (G2P) Conversion☆16Mar 15, 2025Updated last year
- magicspeech competition recipe☆18Jun 29, 2020Updated 6 years ago
- Scaled diffusion transformer for text-to-speech synthesis (DiT + T5Gemma2 conditioning, TorchTitan & Megatron backends, tested up to 1024…☆24Mar 29, 2026Updated 3 months ago
- All-in-one Speech Transcription☆11Jun 5, 2026Updated last month
- This is a subset of the DALI set consisting of 240 polyphonic recordings that is used to benchmark lyrics transcription evaluation.☆12Nov 30, 2021Updated 4 years ago
- Enable RNNLM lattice rescoring with Pytorch [kaldi]☆12Jun 5, 2020Updated 6 years ago
- Custom AppleScript libraries providing a variety of utilities☆18Sep 11, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- M7-TTS: A Mini-Scale Multilingual and Multi-Dialect Text-to-Speech Language Model with Mimi codec and Multi Token Prediction☆20Mar 19, 2026Updated 4 months ago
- audio, NLP, ML with huggingface, nvidia/nemo, speechbrain☆11Sep 4, 2023Updated 2 years ago
- Implementation of the paper "Confidence estimation for attention based sequence to sequence models for speech recognition"☆16May 9, 2021Updated 5 years ago
- Train punctuation and capitalization models for different languages☆26Apr 2, 2022Updated 4 years ago
- Europeanized CosyVoice2 for French & German☆17Mar 30, 2026Updated 3 months ago
- ☆22Jul 22, 2022Updated 3 years ago
- Automatic Speech Recognition tool☆20Aug 5, 2023Updated 2 years ago