☆19Aug 27, 2018Updated 7 years ago
Alternatives and similar repositories for deepsphinx
Users that are interested in deepsphinx are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17Sep 2, 2017Updated 8 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- Desktop like zooming in mobile☆13Mar 8, 2017Updated 9 years ago
- BurrMill core☆22Nov 2, 2021Updated 4 years ago
- text to speech☆10Mar 19, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- A universal phone recognizer that can transcribe speech in 70+ languages into IPA☆31Jun 9, 2026Updated 2 months ago
- Compiled list of links from "Ask HN: Where can I post my startup to get beta users?"☆17Jan 28, 2016Updated 10 years ago
- A toolkit and benchmark for evaluating phonetic capabilities of speech models.☆18Apr 10, 2026Updated 4 months ago
- Free and Open Platform for AI-assisted Computing☆10May 19, 2019Updated 7 years ago
- ODAS: Open embeddeD Audition System☆11Mar 20, 2021Updated 5 years ago
- Text frontend for ESPnet tts recipes☆35Jun 1, 2021Updated 5 years ago
- SIGMORPHON 2020 Shared Task: Grapheme-to-Phoneme, Unsupervised Induction of Morphology, and Typologically Diverse Morphological Inflectio…☆36Apr 25, 2025Updated last year
- ☆10Apr 17, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆16Nov 11, 2024Updated last year
- Kaldi code for doing DNN with tensorflow☆13Feb 8, 2016Updated 10 years ago
- Open-weights voice acting pipeline combining zero-shot voice cloning with natural-language direction. Provide a reference voice (or gener…☆17May 25, 2026Updated 2 months ago
- Speech Recognition Scoring Toolkit☆13Sep 30, 2015Updated 10 years ago
- ☆27Jan 19, 2021Updated 5 years ago
- A live streaming, video hosting and real-time chat application inspired by Twitch☆10Jan 19, 2023Updated 3 years ago
- codebase for the Text-based NP Enrichment (TNE) paper☆19Mar 12, 2024Updated 2 years ago
- A Tensorflow SqueezeNet implementation☆14Oct 1, 2018Updated 7 years ago
- Support material and source code for the model described in : "A Recurrent Encoder-Decoder Approach With Skip-Filtering Connections For M…☆13Sep 19, 2017Updated 8 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Files for the course Offensive Computer Security 2014 (FSU)☆12May 20, 2015Updated 11 years ago
- DUSTED: Spoken-Term Discovery using Discrete Speech Units☆17Oct 2, 2024Updated last year
- Different Tikz templates.☆21Dec 8, 2017Updated 8 years ago
- SWIG bindings for Kaldi I/O, built with Conda☆15Dec 15, 2024Updated last year
- 🔥 Discover trending videos from Reddit and curated YouTube channels – Soon using Next.js. See `dev` branch☆15Updated this week
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated 11 months ago
- Toward Multi Modality Language Model - implementation of GPT-4o/Project Astra☆16Dec 10, 2024Updated last year
- Character level speech recognizer using ctc loss with deep rnns in TensorFlow.☆78Jun 9, 2018Updated 8 years ago
- Repo for the 1-Click VTCLI open-source project☆14Jan 8, 2018Updated 8 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Automatic speech recognition using neural networks☆18Nov 21, 2020Updated 5 years ago
- The Cantonese Wordnet☆15Dec 4, 2023Updated 2 years ago
- Dogs vs. Cats Redux on floydhub☆24Jun 19, 2017Updated 9 years ago
- set of scripts used for Performance Tuning, Capacity Planning and Sizing☆11Mar 5, 2022Updated 4 years ago
- Python wrapper for OpenFST and its extensions from Kaldi. Also support reading/writing ark/scp files☆56Apr 9, 2026Updated 4 months ago
- Code for CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment, Interspeech 2026.☆20Jun 9, 2026Updated 2 months ago
- PyTorch implementation of the ICASSP-24 paper: "Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Superv…☆41Jan 6, 2024Updated 2 years ago