πΈSTT integration examples
β132Sep 23, 2022Updated 3 years ago
Alternatives and similar repositories for STT-examples
Users that are interested in STT-examples are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- πΈSTT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.β2,596Mar 11, 2024Updated 2 years ago
- TTS Client for Coqui TTS serverβ13Jan 7, 2023Updated 3 years ago
- Coqui STT Model Manager - install, manage and try out Coqui STT models from the Model Zooβ26Mar 24, 2023Updated 3 years ago
- π Coqui's machine learning job schedulerβ31Sep 5, 2021Updated 4 years ago
- Coqui Inference Engineβ41Aug 3, 2021Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- πΈTTS recipes for different datasetsβ88Jul 26, 2022Updated 4 years ago
- Evaluation of STT models for german languageβ16Jan 22, 2022Updated 4 years ago
- Open models for Coqui STTβ153May 9, 2023Updated 3 years ago
- Implementation of different noise embeddings for noise aware training of Kaldi acoustic models.β13Feb 13, 2021Updated 5 years ago
- Coqui STT offline engine API for NodeJs developers. With a simple HTTP ASR server.β30Jun 8, 2021Updated 5 years ago
- Linguistic processing for Common Voiceβ59Jan 18, 2024Updated 2 years ago
- A living document for all things Common Voice.β14Jun 24, 2024Updated 2 years ago
- Segment a given audio into utterances using a trained end-to-end ASR model.β75Oct 9, 2020Updated 5 years ago
- πΈ collection of TTS papersβ731Jul 4, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A voice driven 3D chess game for learning Voice AIβ17Jul 6, 2022Updated 4 years ago
- Agile reading group that worksβ13Feb 2, 2022Updated 4 years ago
- Tooling for producing French dataset for Common Voiceβ101Jan 20, 2025Updated last year
- Mozilla Voice Community Playbookβ48May 21, 2024Updated 2 years ago
- π A list of accessible speech corpora for ASR, TTS, and other Speech Technologiesβ1,398Jun 6, 2024Updated 2 years ago
- β55Jan 13, 2023Updated 3 years ago
- Android Speech Recognition Service using Vosk/Kaldi and Mozilla DeepSpeechβ110Jan 19, 2022Updated 4 years ago
- Chinese-ASR built on kaldiβ14Jan 21, 2019Updated 7 years ago
- A library of speech gadgets.β15Oct 15, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- scipts for working with open.bible dataβ26Jan 24, 2022Updated 4 years ago
- The YouTube Text-To-Speech dataset is comprised of waveform audio extracted from YouTube videos alongside their English transcriptionsβ53Apr 1, 2021Updated 5 years ago
- PyTorch speech2text inference script for the NVidia openseq2seq wav2letter model variantβ10Aug 12, 2019Updated 6 years ago
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challengeβ16Mar 26, 2022Updated 4 years ago
- β15Sep 13, 2022Updated 3 years ago
- Fast trigram-indexed regex search for codebases β 2-6x faster than ripgrepβ20Mar 24, 2026Updated 4 months ago
- A VR180 photo viewer that works on a web browser.β12May 18, 2019Updated 7 years ago
- Ultrafast GAN based Vocoder for Text to Speechβ50Jul 16, 2022Updated 4 years ago
- Running Mozilla's implementation of Baidu DeepSpeech on Google Colaboratoryβ16Mar 18, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Script for bundling Common Voice (https://commonvoice.mozilla.org/) clips by languageβ11Apr 13, 2023Updated 3 years ago
- REST api for mozilla deepspeech voice recognition engineβ20Nov 1, 2021Updated 4 years ago
- Run speaker recognition algorithms - Mirrored from https://gitlab.idiap.ch/bob/bob.bio.spearβ19Jun 24, 2023Updated 3 years ago
- Simple tkinter application for recorded voice samples with text promptsβ18Jul 6, 2023Updated 3 years ago
- A proof-of-concept app using KeenASR SDK on Android. WE ARE HIRING: https://keenresearch.com/careers.htmlβ29Jun 30, 2026Updated 3 weeks ago
- π« check your data, before you wreck your modelβ16Aug 11, 2022Updated 3 years ago
- unofficial pytorch implementation of HiFi-GAN with fast MISR.β15Mar 21, 2023Updated 3 years ago