☆19Aug 27, 2018Updated 8 years ago
Alternatives and similar repositories for deepsphinx
Users that are interested in deepsphinx are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17Sep 2, 2017Updated 9 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- BurrMill core☆22Nov 2, 2021Updated 4 years ago
- text to speech☆10Mar 19, 2024Updated 2 years ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- A universal phone recognizer that can transcribe speech in 70+ languages into IPA☆42Updated this week
- Compiled list of links from "Ask HN: Where can I post my startup to get beta users?"☆17Jan 28, 2016Updated 10 years ago
- A toolkit and benchmark for evaluating phonetic capabilities of speech models.☆20Apr 10, 2026Updated 5 months ago
- Text frontend for ESPnet tts recipes☆35Sep 2, 2026Updated 2 weeks ago
- Cantonese Grapheme-to-Phoneme Converter based on GitYCC/g2pW☆15Dec 10, 2024Updated last year
- SIGMORPHON 2020 Shared Task: Grapheme-to-Phoneme, Unsupervised Induction of Morphology, and Typologically Diverse Morphological Inflectio…☆36Apr 25, 2025Updated last year
- ☆10Apr 17, 2024Updated 2 years ago
- Attentional Neural Network that translates text to phones.☆11Jan 25, 2018Updated 8 years ago
- Open-weights voice acting pipeline combining zero-shot voice cloning with natural-language direction. Provide a reference voice (or gener…☆18May 25, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Speech Recognition Scoring Toolkit☆13Sep 30, 2015Updated 10 years ago
- ☆27Jan 19, 2021Updated 5 years ago
- codebase for the Text-based NP Enrichment (TNE) paper☆19Mar 12, 2024Updated 2 years ago
- Support material and source code for the model described in : "A Recurrent Encoder-Decoder Approach With Skip-Filtering Connections For M…☆13Sep 19, 2017Updated 9 years ago
- DUSTED: Spoken-Term Discovery using Discrete Speech Units☆17Oct 2, 2024Updated last year
- Different Tikz templates.☆21Dec 8, 2017Updated 8 years ago
- SWIG bindings for Kaldi I/O, built with Conda☆15Dec 15, 2024Updated last year
- Dynamic Entity Summarization (DynES)☆20May 10, 2019Updated 7 years ago
- Toward Multi Modality Language Model - implementation of GPT-4o/Project Astra☆16Dec 10, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Character level speech recognizer using ctc loss with deep rnns in TensorFlow.☆78Jun 9, 2018Updated 8 years ago
- The Cantonese Wordnet☆15Dec 4, 2023Updated 2 years ago
- Python wrapper for OpenFST and its extensions from Kaldi. Also support reading/writing ark/scp files☆56Apr 9, 2026Updated 5 months ago
- Seq2Seq Chatbot with attention mechanism☆18Apr 27, 2017Updated 9 years ago
- PyTorch implementation of the ICASSP-24 paper: "Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Superv…☆41Jan 6, 2024Updated 2 years ago
- Download and preperation tool for free speech corpora.☆16Apr 28, 2019Updated 7 years ago
- An AR+AR TTS attempt.☆18Jan 13, 2025Updated last year
- Demo for DART, Audio Imagination workshop submission in NeurIPS 2024☆16Apr 22, 2026Updated 5 months ago
- ☆15Jul 24, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Android offline speech recognition natively on PC☆53Dec 13, 2020Updated 5 years ago
- Python implementation of Levenshtein distance and Levenshtein automata matching☆27May 8, 2019Updated 7 years ago
- This repository is built with a focus on practical ways to obtain and work with the audio data of audioset. You can use this repository t…☆18Jun 12, 2025Updated last year
- Code for CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment, Interspeech 2026.☆22Aug 14, 2026Updated last month
- A database of number names for 186 languages, locales, and scripts☆67Mar 3, 2023Updated 3 years ago
- Reference HDL code for the MATRIX Creator's Spartan 6 FPGA☆28Jan 15, 2020Updated 6 years ago
- Parallelized automatic corpus collection for ASR. Forked from https://github.com/EgorLakomkin/KTSpeechCrawler☆23Mar 21, 2021Updated 5 years ago