☆13,057Oct 25, 2025Updated 10 months ago
Alternatives and similar repositories for insanely-fast-whisper
Users that are interested in insanely-fast-whisper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)☆23,922Aug 30, 2026Updated last week
- Faster Whisper transcription with CTranslate2☆25,275Nov 19, 2025Updated 9 months ago
- Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.☆4,115Jan 8, 2025Updated last year
- Port of OpenAI's Whisper model in C/C++☆53,495Updated this week
- Robust Speech Recognition via Large-Scale Weak Supervision☆108,692Aug 31, 2026Updated last week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.☆4,681Apr 3, 2024Updated 2 years ago
- Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper☆5,643Aug 15, 2026Updated 3 weeks ago
- An Open Source text-to-speech system built by inverting Whisper.☆4,646Dec 14, 2025Updated 8 months ago
- Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.☆75,758Updated this week
- WhisperPlus: Faster, Smarter, and More Capable 🚀☆1,957Aug 3, 2026Updated last month
- Open-Source Frontier Voice AI☆53,885Updated this week
- StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models☆6,345Aug 10, 2024Updated 2 years ago
- Open Source AI Platform - AI Chat with advanced features that works with every LLM☆31,960Updated this week
- Instant voice cloning by MIT and MyShell. Audio foundation model.☆37,468Apr 19, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production☆45,987Aug 16, 2024Updated 2 years ago
- SOTA Open Source TTS☆32,589Aug 22, 2026Updated 2 weeks ago
- A coding agent for open models like Kimi K3 and GLM 5.3☆68,263Updated this week
- Foundational Models for State-of-the-Art Speech and Text Translation☆11,852Jul 28, 2026Updated last month
- Platform for stateful agents: AI with advanced memory that can learn and self-improve over time.☆24,641Aug 23, 2026Updated 2 weeks ago
- OCR, layout analysis, reading order, table recognition in 90+ languages☆21,359Updated this week
- Cross-Platform, GPU Accelerated Whisper 🏎️☆1,794Feb 27, 2024Updated 2 years ago
- Neural building blocks for speaker diarization: speech activity detection, speaker change detection, overlapped speech detection, speaker…☆10,515Updated this week
- The fastest Whisper optimization for automatic speech recognition as a command-line interface ⚡️☆409Jun 8, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SoTA open-source TTS☆26,304Jul 21, 2026Updated last month
- 🔊 Text-Prompted Generative Audio Model☆39,263Aug 19, 2024Updated 2 years ago
- Build, run, and manage agent platforms.☆42,085Updated this week
- The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails…☆58,221Updated this week
- DSPy: The framework for programming—not prompting—language models☆37,828Updated this week
- Structured Outputs☆15,759Updated this week
- Vane is an AI-powered answering engine.☆36,663Sep 1, 2026Updated last week
- The Memory Layer for AI Agents - Drop-in memory infrastructure for AI agents and apps. Context that persists. Built for production.☆64,848Updated this week
- Zero-Shot Speech Editing and Text-to-Speech in the Wild☆8,576May 30, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- We write your reusable computer vision tools. 💜☆49,909Updated this week
- Distribute and run LLMs with a single file.☆25,901Updated this week
- 🙌 OpenHands: AI-Driven Development☆86,438Updated this week
- LlamaIndex is the leading document agent and OCR platform☆52,058Updated this week
- Convert PDF to markdown + JSON quickly with high accuracy☆39,570Aug 31, 2026Updated last week
- LLM inference in C/C++☆127,373Updated this week
- Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor…☆23,611Mar 3, 2026Updated 6 months ago