C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more
☆660Sep 22, 2026Updated this week
Alternatives and similar repositories for CrispASR
Users that are interested in CrispASR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- On-device speech-to-text Flutter app powered by CrispASR (ggml / Whisper) — offline, multi-platform, AGPL-3.0.☆62Updated this week
- speech to text gui for different (e.g. Whisper, Voxtral) models and backends, including whisper.cpp, crispasar, mlx-whisper, faster-whisp…☆31Sep 13, 2026Updated last week
- Lightweight text and scans processing: embedding, document processing, OCR, OMR, etc, with inference via ggml in pure C++☆62Updated this week
- Implementation of Qwen3-ASR-0.6B in GGML☆115Jul 28, 2026Updated last month
- C++ port of Microsoft VibeVoice built on ggml☆128Jul 9, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Local AI text-to-speech with voice cloning and voice design, powered by GGML. C++17 port of OmniVoice (k2-fsa/OmniVoice). 646 languages, …☆177Updated this week
- ☆42Jul 13, 2026Updated 2 months ago
- An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, a…☆2,935Updated this week
- CosyVoice inference in C/C++☆55Updated this week
- ONNX speech pipeline library for ASR (diarization, VAD), and TTS☆22Updated this week
- very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust☆397Updated this week
- Parakeet implementation in C++ with ggml☆788Aug 28, 2026Updated 3 weeks ago
- Implementation of Fish Audio S2 Pro model inference in native ggml.☆124May 27, 2026Updated 3 months ago
- ☆22May 2, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- PyTorch -> ONNX☆17Oct 18, 2025Updated 11 months ago
- Standalone C++ inference project for VoxCPM models built on top of ggml.☆94Jul 14, 2026Updated 2 months ago
- Port of Mistral's Voxtral model in C/C++☆36Jun 19, 2026Updated 3 months ago
- Ultra fast and portable Parakeet implementation for on-device inference in C++ using Axiom with MPS+Unified Memory☆306May 4, 2026Updated 4 months ago
- ☆31Feb 14, 2026Updated 7 months ago
- On-device VAD / streaming STT / TTS / diarization in C++17 (ONNX + LiteRT) with a voice-agent pipeline. Linux, Windows, Android.☆89Sep 15, 2026Updated last week
- Portable C++17 implementation of ACE-Step 1.5 AI Music Generator using GGML. Text + lyrics in, stereo 48kHz MP3 or WAV out. Runs on CPU, …☆424Updated this week
- Pure-PyTorch Parakeet TDT inference☆54Mar 10, 2026Updated 6 months ago
- ☆39Mar 30, 2026Updated 5 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- C inference for Qwen3-ASR 0.6b and 1.7b transcriptions models☆615Feb 17, 2026Updated 7 months ago
- ☆241Jul 18, 2026Updated 2 months ago
- A SOTA Industrial-Grade Voice Activity Detection & Audio Event Detection, supporting 100+ languages, outperforming Silero-VAD, TEN-VAD, F…☆535May 6, 2026Updated 4 months ago
- 将 Qwen3-ASR 的 LLM 部分导 出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆236Apr 29, 2026Updated 4 months ago
- Updated voice agent with Nemotron 3.5 ASR☆18Jun 4, 2026Updated 3 months ago
- Local AI text-to-speech with voice cloning and voice design, powered by GGML. C++17 port of Qwen3-TTS (QwenLM/Qwen3-TTS). 10 languages, 2…☆172Updated this week
- A lightweight Python package for Automatic Speech Recognition using ONNX models☆377Aug 16, 2026Updated last month
- An OpenAI-compatible ASR/STT API server powered by Meta's omnilingual-asr model. Supports real-time streaming via WebSocket and batch tra…☆21Jan 2, 2026Updated 8 months ago
- Port of Funasr's Sense-voice model in C/C++☆575Dec 19, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Pure-Rust inference engine for Qwen3-ASR speech recognition models (0.6B & 1.7B) using candle with Metal/CUDA acceleration☆27Mar 17, 2026Updated 6 months ago
- Pure-PyTorch inference for CohereLabs/cohere-transcribe-03-2026 (2B Conformer + Transformer ASR, 14 languages).☆45Apr 29, 2026Updated 4 months ago
- A highly optimized engine for neutts-air model to generate minutes of audio in seconds. Over 200x realtime on modern hardware!☆120Nov 24, 2025Updated 9 months ago
- Cohere Transcribe in Rust☆101May 19, 2026Updated 4 months ago
- Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++☆7,092Updated this week
- 为生产与边缘场景优化的 GPT-SoVITS c++库绑定,ONNX/TensorRT后端.☆19Mar 12, 2026Updated 6 months ago
- Pure C++ implementation of several models for real-time chatting on your computer (CPU & GPU)☆930Sep 7, 2026Updated 2 weeks ago