Google Chrome SODA Offline Speech Recognition command line client
☆171Jan 28, 2025Updated last year
Alternatives and similar repositories for gasr
Users that are interested in gasr are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Google Chrome Text to Speech command line client☆37Jul 16, 2021Updated 5 years ago
- This project aims to research google's offline speech recognition, from several android apps and ideally make them interoperable by repli…☆70May 10, 2020Updated 6 years ago
- Android offline speech recognition natively on PC☆53Dec 13, 2020Updated 5 years ago
- Tensorflow-based wake word detection☆23Sep 17, 2026Updated last week
- This is code for an audio search engine that uses vocal imitations of the desired sound☆38May 16, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Experiment in automatic insertion of timed transcript corrections☆21Oct 31, 2017Updated 8 years ago
- Main docs for the project☆14Nov 4, 2022Updated 3 years ago
- An efficient implementation of RNN-T Prefix Beam Search in C++/CUDA.☆67Jan 7, 2026Updated 8 months ago
- ☆11Aug 11, 2023Updated 3 years ago
- ✨Realtime Voice Changer with 3~ seconds for custom voice in CPU☆23Apr 21, 2026Updated 5 months ago
- Accelerate Whisper tasks such as transcription, by multiprocesing through parallelization☆25Oct 29, 2022Updated 3 years ago
- BurrMill core☆22Nov 2, 2021Updated 4 years ago
- ☆14Aug 19, 2024Updated 2 years ago
- ☆27Jan 19, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Generate audio datasets for training Text-To-Speech models, through smart audio splitting with silence detection, and transcription using…☆30May 27, 2023Updated 3 years ago
- A little Node.JS application that expose an API to do ebook conversion using calibre ebook-convert command.☆11Feb 21, 2022Updated 4 years ago
- Colab notebooks for Next-gen Kaldi☆31Oct 12, 2025Updated 11 months ago
- 流声字幕:基于 ASR 与 LLM 生成、翻译流媒体视频字幕的 Chrome 扩展。 / Chrome extension for ASR + LLM subtitles on streaming videos.☆57Sep 1, 2026Updated 3 weeks ago
- Tiny wrapper around webrtc-audio-processing for noise suppression/auto gain only☆35May 28, 2026Updated 4 months ago
- [ICCV'21] The Right to Talk: An Audio-Visual Transformer Approach☆20Aug 2, 2021Updated 5 years ago
- Convert Korean to Katakana☆13Dec 13, 2023Updated 2 years ago
- Compiled list of links from "Ask HN: Where can I post my startup to get beta users?"☆17Jan 28, 2016Updated 10 years ago
- Python wrapper for OpenFST and its extensions from Kaldi. Also support reading/writing ark/scp files☆56Apr 9, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Decoders from Kaldi using OpenFst☆35Apr 10, 2026Updated 5 months ago
- ☆15Apr 16, 2026Updated 5 months ago
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆13Jul 15, 2024Updated 2 years ago
- Go language bindings for the ggwave C++ library☆14Apr 9, 2025Updated last year
- Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS (E2 TTS) in MLX☆29Oct 15, 2024Updated last year
- finetune the chain model based on cvte open source model without traing any GMM for frame alignment☆12Aug 6, 2020Updated 6 years ago
- repo of files pertaining to realtime, offline translations using whisper realtime and argos translate. This repo is marked Creative Commo…☆19May 20, 2025Updated last year
- ☆15Oct 25, 2024Updated last year
- Sequence to sequence model for Arabic punctuation prediction.☆12Feb 13, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ODAS: Open embeddeD Audition System☆11Mar 20, 2021Updated 5 years ago
- DiTTo-TTS: Diffusion Transformers for Scalable Text-to-Speech without Domain-Specific Factors☆39Feb 11, 2025Updated last year
- Assistance component base for Dicio assistant components☆14Apr 23, 2026Updated 5 months ago
- Robust Speech Recognition via Large-Scale Weak Supervision☆91Aug 28, 2023Updated 3 years ago
- Text frontend for ESPnet tts recipes☆35Sep 2, 2026Updated 3 weeks ago
- An experiment of trying out whisper.cpp for real-time speech-to-text☆19Dec 25, 2022Updated 3 years ago
- ☆22Jun 30, 2021Updated 5 years ago