C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more
☆539Aug 13, 2026Updated this week
Alternatives and similar repositories for CrispASR
Users that are interested in CrispASR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- On-device speech-to-text Flutter app powered by CrispASR (ggml / Whisper) — offline, multi-platform, AGPL-3.0.☆46Aug 5, 2026Updated last week
- speech to text gui for different (e.g. Whisper, Voxtral) models and backends, including whisper.cpp, crispasar, mlx-whisper, faster-whisp…☆30Updated this week
- Lightweight text and scans processing: embedding, document processing, OCR, OMR, etc, with inference via ggml in pure C++☆47Updated this week
- Implementation of Qwen3-ASR-0.6B in GGML☆108Jul 28, 2026Updated 2 weeks ago
- C++ port of Microsoft VibeVoice built on ggml☆118Jul 9, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Local AI text-to-speech with voice cloning and voice design, powered by GGML. C++17 port of OmniVoice (k2-fsa/OmniVoice). 646 languages, …☆156Jul 21, 2026Updated 3 weeks ago
- ☆39Jul 13, 2026Updated last month
- An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, a…☆1,304Updated this week
- CosyVoice inference in C/C++☆44Updated this week
- ONNX speech pipeline library for ASR, diarization, VAD, and denoising☆20Jun 14, 2026Updated 2 months ago
- very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust☆383Jul 28, 2026Updated 2 weeks ago
- Parakeet implementation in C++ with ggml☆761Aug 1, 2026Updated last week
- ☆21May 2, 2026Updated 3 months ago
- PyTorch -> ONNX☆17Oct 18, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Standalone C++ inference project for VoxCPM models built on top of ggml.☆89Jul 14, 2026Updated last month
- Port of Mistral's Voxtral model in C/C++☆33Jun 19, 2026Updated last month
- Ultra fast and portable Parakeet implementation for on-device inference in C++ using Axiom with MPS+Unified Memory☆301May 4, 2026Updated 3 months ago
- On-device VAD / streaming STT / TTS / diarization in C++17 (ONNX + LiteRT) with a voice-agent pipeline. Linux, Windows, Android.☆67Aug 5, 2026Updated last week
- Portable C++17 implementation of ACE-Step 1.5 AI Music Generator using GGML. Text + lyrics in, stereo 48kHz MP3 or WAV out. Runs on CPU, …☆394Updated this week
- Pure-PyTorch Parakeet TDT inference☆52Mar 10, 2026Updated 5 months ago
- C inference for Qwen3-ASR 0.6b and 1.7b transcriptions models☆592Feb 17, 2026Updated 5 months ago
- ☆227Jul 18, 2026Updated 3 weeks ago
- ☆28Feb 14, 2026Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A SOTA Industrial-Grade Voice Activity Detection & Audio Event Detection, supporting 100+ languages, outperforming Silero-VAD, TEN-VAD, F…☆501May 6, 2026Updated 3 months ago
- 将 Qwen3-ASR 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆212Apr 29, 2026Updated 3 months ago
- Local AI text-to-speech with voice cloning and voice design, powered by GGML. C++17 port of Qwen3-TTS (QwenLM/Qwen3-TTS). 10 languages, 2…☆130Aug 7, 2026Updated last week
- A lightweight Python package for Automatic Speech Recognition using ONNX models☆361Aug 4, 2026Updated last week
- ☆37Mar 30, 2026Updated 4 months ago
- An OpenAI-compatible ASR/STT API server powered by Meta's omnilingual-asr model. Supports real-time streaming via WebSocket and batch tra…☆18Jan 2, 2026Updated 7 months ago
- 为生产与边缘场景优化的 GPT-SoVITS c++库绑定,ONNX/TensorRT后端.☆17Mar 12, 2026Updated 5 months ago
- Pure-PyTorch inference for CohereLabs/cohere-transcribe-03-2026 (2B Conformer + Transformer ASR, 14 languages).☆41Apr 29, 2026Updated 3 months ago
- A highly optimized engine for neutts-air model to generate minutes of audio in seconds. Over 200x realtime on modern hardware!☆119Nov 24, 2025Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Port of Funasr's Sense-voice model in C/C++☆571Dec 19, 2025Updated 7 months ago
- sa3 implemented in ggml for cross-platform and embedded applications☆18Updated this week
- Cohere Transcribe in Rust☆97May 19, 2026Updated 2 months ago
- Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++☆6,736Updated this week
- Pure C++ implementation of several models for real-time chatting on your computer (CPU & GPU)☆915Aug 4, 2026Updated last week
- Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp…☆1,474Jul 24, 2026Updated 3 weeks ago
- Rust implementation of Qwen3-ASR automatic speech recognition☆247Mar 28, 2026Updated 4 months ago