An automatic speech recognition API
☆86Sep 8, 2026Updated 2 weeks ago
Alternatives and similar repositories for linto-stt
Users that are interested in linto-stt are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Build a LinTO OS Image which boots on Raspberry Pi3☆14Jul 8, 2020Updated 6 years ago
- Open source transcription, live subtitling and meeting summarization. Web app, API and SDKs of the LinTO platform.☆59Updated this week
- LinTO platform services stack deployment tool for Docker Swarm cluster☆16Feb 27, 2024Updated 2 years ago
- Tools for speech processing, keyword spotting☆16Mar 11, 2020Updated 6 years ago
- Lists of conversational datasets☆21Sep 5, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Speaker diarization service☆27Jul 2, 2026Updated 2 months ago
- GUI Tool to create, manage and test Keyword Spotting models using TF 2.0☆13Feb 1, 2021Updated 5 years ago
- This repository provides data and code for "Vox Populi, Vox DIY: Benchmark Dataset for Crowdsourced Audio Transcription" paper.☆16Jul 22, 2021Updated 5 years ago
- ☆13Aug 7, 2021Updated 5 years ago
- phone inventory library☆18May 15, 2023Updated 3 years ago
- enhan(t) is an open source toolkit which enables you to enhance the web experience of existing video conferencing solutions like Zoom, MS…☆15Apr 28, 2022Updated 4 years ago
- Thai Grapheme to Phoneme (G2P) Wiktionary Corpus☆13Jul 25, 2022Updated 4 years ago
- TTS for Singlish using Tacotron2, the IMDA corpus, and Pachyderm.☆11Jan 11, 2020Updated 6 years ago
- ☆17Jun 30, 2020Updated 6 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- finetune the chain model based on cvte open source model without traing any GMM for frame alignment☆12Aug 6, 2020Updated 6 years ago
- Next-generation Punkt sentence boundary detection with zero dependencies☆32Nov 18, 2025Updated 10 months ago
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.☆13Feb 13, 2021Updated 5 years ago
- ☆17Apr 14, 2023Updated 3 years ago
- ☆11Sep 5, 2025Updated last year
- Model for recasing and repunctuating ASR transcripts☆142Apr 10, 2024Updated 2 years ago
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆35Sep 25, 2025Updated 11 months ago
- radiomixer☆14Feb 16, 2022Updated 4 years ago
- ☆20Jun 16, 2026Updated 3 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Character-level conversion between Hebrew text and Latin transliteration using deep learning - a demonstration of seq2seq training.☆15Jun 27, 2023Updated 3 years ago
- DEPRECATED - A webapp for collecting speech samples for voice recognition testing and training☆20May 23, 2019Updated 7 years ago
- magicspeech competition recipe☆18Jun 29, 2020Updated 6 years ago
- brainless concatenative text to speech☆16May 11, 2021Updated 5 years ago
- Scaled diffusion transformer for text-to-speech synthesis (DiT + T5Gemma2 conditioning, TorchTitan & Megatron backends, tested up to 1024…☆25Mar 29, 2026Updated 5 months ago
- All-in-one Speech Transcription☆11Jun 5, 2026Updated 3 months ago
- Enable RNNLM lattice rescoring with Pytorch [kaldi]☆12Jun 5, 2020Updated 6 years ago
- M7-TTS: A Mini-Scale Multilingual and Multi-Dialect Text-to-Speech Language Model with Mimi codec and Multi Token Prediction☆20Mar 19, 2026Updated 6 months ago
- audio, NLP, ML with huggingface, nvidia/nemo, speechbrain☆11Sep 4, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆13Feb 24, 2025Updated last year
- Implementation of the paper "Confidence estimation for attention based sequence to sequence models for speech recognition"☆16May 9, 2021Updated 5 years ago
- Train punctuation and capitalization models for different languages☆26Apr 2, 2022Updated 4 years ago
- ☆23Jul 22, 2022Updated 4 years ago
- An upgrade framework for train and validate compare with icefall using Lightning.☆16Mar 26, 2025Updated last year
- A Unity implementation of DeepSpeech is an open source embedded (offline, on-device) speech-to-text engine which can run in real time on …☆26Sep 22, 2022Updated 4 years ago
- homepage of DreamActor-M1☆64Jun 26, 2025Updated last year