☆45Apr 28, 2026Updated 2 months ago
Alternatives and similar repositories for granite-speech-models
Users that are interested in granite-speech-models are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Oct 24, 2025Updated 8 months ago
- Ultra-Sortformer for Scalable Speaker Diarization☆27Apr 9, 2026Updated 3 months ago
- FMS Model Optimizer is a framework for developing reduced precision neural network models.☆21Jun 24, 2026Updated 3 weeks ago
- A simple script to prepare dataset for training with TTS Tortoise model via https://git.ecker.tech/mrq/ai-voice-cloning☆12Jan 12, 2024Updated 2 years ago
- ☆30Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- High-performance, semantic turn detection for conversational AI☆44Oct 1, 2025Updated 9 months ago
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequ…☆31Sep 20, 2025Updated 10 months ago
- Glow-TTS with Stochastic Duration Predictor and Stochastic Pitch Predictor☆19Jun 5, 2023Updated 3 years ago
- The accompanying code for "Exploring the limits of decoder-only models trained on public speech recognition corpora" (Ankit Gupta, George…☆21Oct 11, 2024Updated last year
- ☆46Apr 29, 2026Updated 2 months ago
- We introduce the LLAMA1 Test Set, a comprehensive open-domain world knowledge QA dataset for evaluating question-answering systems. We pr…☆23Mar 14, 2024Updated 2 years ago
- Speech Resynthesis and Language Modeling☆27Jun 11, 2025Updated last year
- ☆33Oct 28, 2025Updated 8 months ago
- 🚀 Collection of libraries used with fms-hf-tuning to accelerate fine-tuning and training of large models.☆14Jan 30, 2026Updated 5 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆22Jan 27, 2026Updated 5 months ago
- Variable Bitrate Residual Vector Quantization for Audio Coding☆54May 1, 2025Updated last year
- ☆18Nov 19, 2025Updated 8 months ago
- A lightweight library for normalizing speech transcripts before computing WER☆27Jul 14, 2026Updated last week
- FlowMirror-HydraVox — A natively accelerated multi-head autoregressive TTS system derived from CosyVoice 3.0. It predicts multiple tokens…☆49Feb 17, 2026Updated 5 months ago
- pytorch model for contexless-phoneme prediction from speech audio☆32Oct 30, 2025Updated 8 months ago
- [ACL 2026 Findings] Affectron: Emotional Speech Synthesis with Affective and Contextually Aligned Nonverbal Vocalizations☆20Updated this week
- ☆35Sep 6, 2025Updated 10 months ago
- ☆28Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Speech recognition module for Python, supporting several engines and APIs, online and offline.☆13Mar 9, 2022Updated 4 years ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- Java Bindings for the C++ library DeepSpeech☆10Jun 4, 2020Updated 6 years ago
- Text-To-Speech for NotebookLM☆39Jul 20, 2025Updated last year
- ☆10May 16, 2023Updated 3 years ago
- ☆13May 1, 2026Updated 2 months ago
- ☆59Apr 1, 2026Updated 3 months ago
- ☆19Aug 23, 2024Updated last year
- ☆104Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A Weakly Supervised Forced Alignment for disluent speech☆15Nov 12, 2023Updated 2 years ago
- ☆13Oct 27, 2021Updated 4 years ago
- Fully local voice interface for Claude Code on Apple Silicon. Parakeet STT + Kokoro TTS + SmartTurn EOU + dual VAD.☆30Mar 24, 2026Updated 3 months ago
- superfast text to speech in any voice☆62Feb 16, 2026Updated 5 months ago
- Code for Latent Speech-Text Transformer (LST)☆35Mar 12, 2026Updated 4 months ago
- ⚠️ Deprecated: This library's functionality has been rolled into Mellea (https://github.com/generative-computing/mellea)☆56Jun 30, 2026Updated 3 weeks ago
- A JAX library for building lattice-based speech transducer models☆48Jul 2, 2026Updated 2 weeks ago