Open models for Coqui STT
☆153May 9, 2023Updated 3 years ago
Alternatives and similar repositories for STT-models
Users that are interested in STT-models are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- TTS Client for Coqui TTS server☆13Jan 7, 2023Updated 3 years ago
- A Text-To-Speech Model Developed Using 🐸STT☆13Jun 22, 2022Updated 4 years ago
- Linguistic processing for Common Voice☆59Jan 18, 2024Updated 2 years ago
- 🫠 check your data, before you wreck your model☆16Aug 11, 2022Updated 3 years ago
- 🐸STT integration examples☆132Sep 23, 2022Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.☆2,596Mar 11, 2024Updated 2 years ago
- Coqui Inference Engine☆41Aug 3, 2021Updated 4 years ago
- 🐸TTS recipes for different datasets☆88Jul 26, 2022Updated 4 years ago
- 🐸 - A general purpose model trainer, as flexible as it gets☆233Mar 7, 2024Updated 2 years ago
- Coqui STT Model Manager - install, manage and try out Coqui STT models from the Model Zoo☆26Mar 24, 2023Updated 3 years ago
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.☆13Feb 13, 2021Updated 5 years ago
- Using YouTube to prepare a speech recognition dataset for any language☆10Mar 30, 2021Updated 5 years ago
- Coqui STT offline engine API for NodeJs developers. With a simple HTTP ASR server.☆30Jun 8, 2021Updated 5 years ago
- This is an ASR corpus for Bemba language. It contains read speech from diverse publicly available Bemba sources; Literature Books, Radio/…☆41Jul 31, 2025Updated 11 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This project builds a custom question answering chatbot using Langchain and Google Gemini Language Model (LLM). It fine-tunes industrial …☆14Apr 2, 2024Updated 2 years ago
- 🐸Coqui Dialogue Audio Pack contains more than 2000 audio files of synthetic human voices over dialogue created specifically for video ga…☆46Mar 7, 2023Updated 3 years ago
- Turkish Speech Recognition using Facebook's Wav2vec 2.0 models☆33Feb 7, 2022Updated 4 years ago
- 👄🇧🇷 Alinhamento fonético forçado em Português Brasileiro☆13Jul 18, 2025Updated last year
- A library of speech gadgets.☆15Oct 15, 2022Updated 3 years ago
- This repository provides data and code for "Vox Populi, Vox DIY: Benchmark Dataset for Crowdsourced Audio Transcription" paper.☆16Jul 22, 2021Updated 5 years ago
- Java Bindings for the C++ library DeepSpeech☆10Jun 4, 2020Updated 6 years ago
- On-device voice activity detection (VAD) powered by deep learning☆266Updated this week
- ☆15Dec 12, 2019Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This app is intended to automatically create a corpus for ASR systems using pseudo-labeling.☆27Feb 15, 2024Updated 2 years ago
- ☆13Oct 27, 2021Updated 4 years ago
- ☆10Mar 8, 2023Updated 3 years ago
- Repository for multilingual speech data resources for native languages of Zambia.☆22Oct 9, 2024Updated last year
- Source code for 'Transfer Learning for Speech Recognition on a Budget' published at ACL 2017☆46May 30, 2017Updated 9 years ago
- Lite Voice Terminal, an "offline smart speaker" solution powered by on-premise ASR server (vosk API / kaldi engine)☆19Feb 29, 2024Updated 2 years ago
- Scraping Wikipedia for fair use sentences☆54Jan 25, 2024Updated 2 years ago
- NeMo: a toolkit for conversational AI☆10Jan 18, 2023Updated 3 years ago
- Source code for ASRU 2019 paper "Adapting Pretrained Transformer to Lattices for Spoken Language Understanding"☆10Jul 8, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A Flask web application to calculate and plot drug concentration over time.☆15Jan 1, 2019Updated 7 years ago
- REST api for mozilla deepspeech voice recognition engine☆20Nov 1, 2021Updated 4 years ago
- simple to use, pretrained/training-less models for speaker diarization☆22Aug 23, 2023Updated 2 years ago
- [APSIPA'22] Exploring Speaker Age Estimation on Different Self-Supervised Learning Models☆14Oct 19, 2022Updated 3 years ago
- DeepSpeech based forced alignment tool☆239Dec 12, 2020Updated 5 years ago
- A Simple Flask App to interact with your Machine Translation Model☆13Feb 26, 2020Updated 6 years ago
- Create an LJSpeech structured voice dataset on wave input☆38Sep 28, 2024Updated last year