Google Chrome SODA Offline Speech Recognition command line client
☆170Jan 28, 2025Updated last year
Alternatives and similar repositories for gasr
Users that are interested in gasr are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Android offline speech recognition natively on PC☆53Dec 13, 2020Updated 5 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- Tensorflow-based wake word detection☆22Jun 22, 2026Updated 2 months ago
- Kaldi code for doing DNN with tensorflow☆13Feb 8, 2016Updated 10 years ago
- This is code for an audio search engine that uses vocal imitations of the desired sound☆38May 16, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Experiment in automatic insertion of timed transcript corrections☆21Oct 31, 2017Updated 8 years ago
- ☆19Aug 27, 2018Updated 8 years ago
- An efficient implementation of RNN-T Prefix Beam Search in C++/CUDA.☆67Jan 7, 2026Updated 8 months ago
- ☆11Aug 11, 2023Updated 3 years ago
- Accelerate Whisper tasks such as transcription, by multiprocesing through parallelization☆25Oct 29, 2022Updated 3 years ago
- BurrMill core☆22Nov 2, 2021Updated 4 years ago
- ☆14Aug 19, 2024Updated 2 years ago
- Distributable shell scripts with dependencies☆11Dec 24, 2016Updated 9 years ago
- ☆27Jan 19, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆14Aug 1, 2025Updated last year
- Generate audio datasets for training Text-To-Speech models, through smart audio splitting with silence detection, and transcription using…☆30May 27, 2023Updated 3 years ago
- Arabic Grapheme-to-Phoneme (G2P) Conversion☆15Mar 15, 2025Updated last year
- DDPM-based Pitch Generation and Pitch Controllable Voice Synthesis.☆55Sep 25, 2023Updated 2 years ago
- A little Node.JS application that expose an API to do ebook conversion using calibre ebook-convert command.☆11Feb 21, 2022Updated 4 years ago
- Colab notebooks for Next-gen Kaldi☆31Oct 12, 2025Updated 10 months ago
- A sample Android app using [whisper.cpp](https://github.com/ggerganov/whisper.cpp/) to do voice-to-text transcriptions.☆64Sep 6, 2023Updated 3 years ago
- Tiny wrapper around webrtc-audio-processing for noise suppression/auto gain only☆34May 28, 2026Updated 3 months ago
- [ICCV'21] The Right to Talk: An Audio-Visual Transformer Approach☆20Aug 2, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Convert Korean to Katakana☆13Dec 13, 2023Updated 2 years ago
- Python wrapper for OpenFST and its extensions from Kaldi. Also support reading/writing ark/scp files☆56Apr 9, 2026Updated 4 months ago
- Decoders from Kaldi using OpenFst☆35Apr 10, 2026Updated 4 months ago
- ☆15Apr 16, 2026Updated 4 months ago
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆13Jul 15, 2024Updated 2 years ago
- Go language bindings for the ggwave C++ library☆14Apr 9, 2025Updated last year
- repo of files pertaining to realtime, offline translations using whisper realtime and argos translate. This repo is marked Creative Commo…☆19May 20, 2025Updated last year
- Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS (E2 TTS) in MLX☆29Oct 15, 2024Updated last year
- Sequence to sequence model for Arabic punctuation prediction.☆12Feb 13, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Solfège learning for Android☆11Nov 7, 2020Updated 5 years ago
- ODAS: Open embeddeD Audition System☆11Mar 20, 2021Updated 5 years ago
- DiTTo-TTS: Diffusion Transformers for Scalable Text-to-Speech without Domain-Specific Factors☆39Feb 11, 2025Updated last year
- Assistance component base for Dicio assistant components☆14Apr 23, 2026Updated 4 months ago
- Text frontend for ESPnet tts recipes☆35Updated this week
- ☆22Jun 30, 2021Updated 5 years ago
- Semi-supervised Learning for Multi-speaker Text-to-speech Synthesis Using Discrete Speech Representation☆39Jul 16, 2020Updated 6 years ago