☆49Sep 16, 2026Updated this week
Alternatives and similar repositories for granite-speech-models
Users that are interested in granite-speech-models are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Sep 11, 2026Updated last week
- Ultra-Sortformer for Scalable Speaker Diarization☆30Apr 9, 2026Updated 5 months ago
- FMS Model Optimizer is a framework for developing reduced precision neural network models.☆21Sep 9, 2026Updated last week
- A simple script to prepare dataset for training with TTS Tortoise model via https://git.ecker.tech/mrq/ai-voice-cloning☆13Jan 12, 2024Updated 2 years ago
- ☆35Aug 6, 2026Updated last month
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- High-performance, semantic turn detection for conversational AI☆45Oct 1, 2025Updated 11 months ago
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequ…☆31Sep 20, 2025Updated last year
- Glow-TTS with Stochastic Duration Predictor and Stochastic Pitch Predictor☆19Jun 5, 2023Updated 3 years ago
- The accompanying code for "Exploring the limits of decoder-only models trained on public speech recognition corpora" (Ankit Gupta, George…☆21Oct 11, 2024Updated last year
- ☆53Apr 29, 2026Updated 4 months ago
- Speech Resynthesis and Language Modeling☆27Jun 11, 2025Updated last year
- We introduce the LLAMA1 Test Set, a comprehensive open-domain world knowledge QA dataset for evaluating question-answering systems. We pr…☆23Mar 14, 2024Updated 2 years ago
- 🚀 Collection of libraries used with fms-hf-tuning to accelerate fine-tuning and training of large models.☆14Sep 8, 2026Updated last week
- ☆36Oct 28, 2025Updated 10 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Variable Bitrate Residual Vector Quantization for Audio Coding☆56May 1, 2025Updated last year
- ☆27Jan 27, 2026Updated 7 months ago
- ☆18Nov 19, 2025Updated 10 months ago
- A lightweight library for normalizing speech transcripts before computing WER☆29Updated this week
- FlowMirror-HydraVox — A natively accelerated multi-head autoregressive TTS system derived from CosyVoice 3.0. It predicts multiple tokens…☆49Feb 17, 2026Updated 7 months ago
- pytorch model for contexless-phoneme prediction from speech audio☆32Oct 30, 2025Updated 10 months ago
- [ACL 2026 Findings] Affectron: Emotional Speech Synthesis with Affective and Contextually Aligned Nonverbal Vocalizations☆21Jul 16, 2026Updated 2 months ago
- Wav2vec2 Large XLSR 53 fine-tuned for Malayalam☆11Sep 7, 2021Updated 5 years ago
- ☆32Aug 18, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆35Sep 6, 2025Updated last year
- Speech recognition module for Python, supporting several engines and APIs, online and offline.☆13Mar 9, 2022Updated 4 years ago
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- Java Bindings for the C++ library DeepSpeech☆10Jun 4, 2020Updated 6 years ago
- Text-To-Speech for NotebookLM☆39Jul 20, 2025Updated last year
- ☆13May 1, 2026Updated 4 months ago
- ☆63Apr 1, 2026Updated 5 months ago
- ☆20Aug 23, 2024Updated 2 years ago
- ☆124Aug 7, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A Weakly Supervised Forced Alignment for disluent speech☆15Nov 12, 2023Updated 2 years ago
- ☆13Oct 27, 2021Updated 4 years ago
- Fully local voice interface for Claude Code on Apple Silicon. Parakeet STT + Kokoro TTS + SmartTurn EOU + dual VAD.☆35Mar 24, 2026Updated 5 months ago
- superfast text to speech in any voice☆63Feb 16, 2026Updated 7 months ago
- Code for Latent Speech-Text Transformer (LST)☆35Mar 12, 2026Updated 6 months ago
- ⚠️ Deprecated: This library's functionality has been rolled into Mellea (https://github.com/generative-computing/mellea)☆56Jun 30, 2026Updated 2 months ago
- A JAX library for building lattice-based speech transducer models☆48Sep 10, 2026Updated last week