Self-hosted realtime and offline ASR with SOTA open models, speaker diarization and an OpenAI-compatible API. GPU, CPU, macOS Apple Silicon.
☆353Sep 28, 2026Updated this week
Alternatives and similar repositories for AsrServe
Users that are interested in AsrServe are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Fun-ASR-Nano-2512官方发布的仓库内容有点多,部署起来坑也比较多,本项目提供一个简化的部署方案。☆151Dec 26, 2025Updated 9 months ago
- qwen3 asr server for openai compatible API☆44Mar 11, 2026Updated 6 months ago
- ☆87Updated this week
- Fun-ASR is an end-to-end speech recognition large model launched by Tongyi Lab.☆112Jul 7, 2026Updated 2 months ago
- Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music…☆3,634Jun 26, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 开箱即用的本地私有化部署语音服务,快速搭建Qwen3ASR/FunASR与Qwen3TTS/CosyVoice后端☆164Jul 6, 2026Updated 2 months ago
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆18Jun 27, 2026Updated 3 months ago
- Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp…☆1,556Sep 10, 2026Updated 3 weeks ago
- 一个基于 Sherpa-ONNX 的高性能语音识别服务,支持实时VAD(语音活动检测)、多语言语音识别和声纹识别功能。☆116Jan 4, 2026Updated 8 months ago
- livekit中文插件☆56Sep 14, 2026Updated 2 weeks ago
- 这是基于FunASR实现的区分说话人语音识别API | This is a speaker-diarization-based speech recognition API implemented using FunASR.☆28Jun 16, 2026Updated 3 months ago
- low-latency realtime ASR based on FireRedASR☆61Jul 8, 2025Updated last year
- 最棒的的ASR后处理热词方案,基于音素编辑距离,实现热词替换。☆70Jun 10, 2026Updated 3 months ago
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆20,568Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Fun-ASR-Nano-2512的Docker版本☆18Jan 10, 2026Updated 8 months ago
- Official Python toolkit for the Qwen3-ASR API. Parallel high‑throughput calls, robust long‑audio transcription, multi‑sample‑rate support…☆1,012Feb 5, 2026Updated 7 months ago
- A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/…☆696Jun 2, 2026Updated 4 months ago
- Python runtime for WeTextProcessing (does not depend on Pynini)☆61Sep 9, 2026Updated 3 weeks ago
- 基于 FunASR SenseVoice 模型的实时语音识别 服务,支持说话人识别、音频降噪、ASR 错误修正等高级功能。☆22Updated this week
- Lightweight streaming Voice Activity Detection (VAD) tool with ONNX runtime☆24Mar 18, 2026Updated 6 months ago
- Pseudo Streaming SenseVoice with Hotwords☆472Sep 23, 2026Updated last week
- Local-first meeting transcription — audio & transcripts never leave your machine. Live captions + upload diarization + voice matching.☆133Aug 3, 2026Updated 2 months ago
- A small and simple example showing how to run Qwen3-ASR with ONNX Runtime.☆40Apr 8, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- FastAPI to serve Qwen-ASR with streaming support. Tested. Benchmarked. Flash Attention 2. Fast & Stable.☆16Jun 24, 2026Updated 3 months ago
- An OpenAI API compatible Automatic Speech Recognition (ASR) server that supports both offline and streaming transcription.☆62Aug 30, 2026Updated last month
- Dolphin is a multilingual, multitask ASR model jointly trained by DataoceanAI and Tsinghua University.☆793Jun 11, 2026Updated 3 months ago
- Implementation of Qwen3-ASR-0.6B in GGML☆114Jul 28, 2026Updated 2 months ago
- X-ASR is a series of automatic speech recognition models based on the icefall framework, focusing on streaming ASR and low-latency deploy…☆186Jul 29, 2026Updated 2 months ago
- [ACL 2026 Main] Open-Ended Speaking Style Modeling via Fine-Grained and Multi-Granular Contrastive Language-Speech Pre-training☆107Apr 6, 2026Updated 5 months ago
- Utilizes ONNX Runtime to transcribe audio into text.☆89Aug 26, 2026Updated last month
- High-performance OCR microservice based on PaddleOCR-VL-0.9B (PaddleOCR-VL-1.5-0.9B) with MinerU-compatible API☆37Jan 30, 2026Updated 8 months ago
- GLM-ASR-Nano: A robust, open-source speech recognition model with 1.5B parameters☆854Mar 6, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A lightweight demo of FunASR-Nano using ONNX runtime.☆88Feb 25, 2026Updated 7 months ago
- Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime…☆15,079Sep 22, 2026Updated last week
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- Streaming Text to Speech Web UI☆22May 6, 2024Updated 2 years ago
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- 将 Qwen3-ASR 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆236Apr 29, 2026Updated 5 months ago
- Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR be…☆1,998Feb 25, 2026Updated 7 months ago