☆45Apr 28, 2026Updated 4 months ago
Alternatives and similar repositories for granite-speech-models
Users that are interested in granite-speech-models are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Oct 24, 2025Updated 10 months ago
- Ultra-Sortformer for Scalable Speaker Diarization☆28Apr 9, 2026Updated 4 months ago
- FMS Model Optimizer is a framework for developing reduced precision neural network models.☆21Jun 24, 2026Updated 2 months ago
- A simple script to prepare dataset for training with TTS Tortoise model via https://git.ecker.tech/mrq/ai-voice-cloning☆12Jan 12, 2024Updated 2 years ago
- ☆33Aug 6, 2026Updated 3 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- High-performance, semantic turn detection for conversational AI☆46Oct 1, 2025Updated 10 months ago
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequ…☆31Sep 20, 2025Updated 11 months ago
- Glow-TTS with Stochastic Duration Predictor and Stochastic Pitch Predictor☆19Jun 5, 2023Updated 3 years ago
- The accompanying code for "Exploring the limits of decoder-only models trained on public speech recognition corpora" (Ankit Gupta, George…☆21Oct 11, 2024Updated last year
- ☆50Apr 29, 2026Updated 4 months ago
- Speech Resynthesis and Language Modeling☆27Jun 11, 2025Updated last year
- We introduce the LLAMA1 Test Set, a comprehensive open-domain world knowledge QA dataset for evaluating question-answering systems. We pr…☆23Mar 14, 2024Updated 2 years ago
- ☆34Oct 28, 2025Updated 10 months ago
- 🚀 Collection of libraries used with fms-hf-tuning to accelerate fine-tuning and training of large models.☆14Jan 30, 2026Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆22Jan 27, 2026Updated 7 months ago
- Variable Bitrate Residual Vector Quantization for Audio Coding☆55May 1, 2025Updated last year
- ☆18Nov 19, 2025Updated 9 months ago
- A lightweight library for normalizing speech transcripts before computing WER☆29Jul 14, 2026Updated last month
- FlowMirror-HydraVox — A natively accelerated multi-head autoregressive TTS system derived from CosyVoice 3.0. It predicts multiple tokens…☆49Feb 17, 2026Updated 6 months ago
- pytorch model for contexless-phoneme prediction from speech audio☆32Oct 30, 2025Updated 10 months ago
- [ACL 2026 Findings] Affectron: Emotional Speech Synthesis with Affective and Contextually Aligned Nonverbal Vocalizations☆21Jul 16, 2026Updated last month
- ☆35Sep 6, 2025Updated 11 months ago
- ☆31Aug 18, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Speech recognition module for Python, supporting several engines and APIs, online and offline.☆13Mar 9, 2022Updated 4 years ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- Java Bindings for the C++ library DeepSpeech☆10Jun 4, 2020Updated 6 years ago
- Text-To-Speech for NotebookLM☆39Jul 20, 2025Updated last year
- ☆13May 1, 2026Updated 3 months ago
- ☆62Apr 1, 2026Updated 4 months ago
- ☆20Aug 23, 2024Updated 2 years ago
- ☆115Aug 7, 2026Updated 3 weeks ago
- A Weakly Supervised Forced Alignment for disluent speech☆15Nov 12, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆13Oct 27, 2021Updated 4 years ago
- Fully local voice interface for Claude Code on Apple Silicon. Parakeet STT + Kokoro TTS + SmartTurn EOU + dual VAD.☆33Mar 24, 2026Updated 5 months ago
- superfast text to speech in any voice☆63Feb 16, 2026Updated 6 months ago
- Code for Latent Speech-Text Transformer (LST)☆35Mar 12, 2026Updated 5 months ago
- ⚠️ Deprecated: This library's functionality has been rolled into Mellea (https://github.com/generative-computing/mellea)☆56Jun 30, 2026Updated 2 months ago
- A JAX library for building lattice-based speech transducer models☆48Jul 29, 2026Updated last month
- Multilingual and code-switching ASR challenges for low resource Indian languages.☆23Jul 26, 2021Updated 5 years ago