All in one Qwen3-ASR Server, compatible with OpenAI API
☆320Jul 14, 2026Updated last week
Alternatives and similar repositories for qwen3-asr
Users that are interested in qwen3-asr are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Fun-ASR-Nano-2512官方发布的仓库内容有点多,部署起来坑也比较多,本项目提供一个简化的部署方案。☆151Dec 26, 2025Updated 6 months ago
- qwen3 asr server for openai compatible API☆42Mar 11, 2026Updated 4 months ago
- ☆85Mar 9, 2026Updated 4 months ago
- Fun-ASR is an end-to-end speech recognition large model launched by Tongyi Lab.☆107Jul 7, 2026Updated 2 weeks ago
- Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music…☆3,214Jun 26, 2026Updated 3 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 开箱即用的本地私有化部署语音服务,快速搭建Qwen3ASR/FunASR与Qwen3TTS/CosyVoice后端☆153Jul 6, 2026Updated 2 weeks ago
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆15Jun 27, 2026Updated 3 weeks ago
- Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp…☆1,425Updated this week
- 一个基于 Sherpa-ONNX 的高性能语音识别服务,支持实时VAD(语音活动检测)、多语言语音识别和声纹识别功能。☆115Jan 4, 2026Updated 6 months ago
- livekit agent plugins☆47Apr 21, 2026Updated 3 months ago
- 这是基于FunASR实现的区分说话人语音识别API | This is a speaker-diarization-based speech recognition API implemented using FunASR.☆27Jun 16, 2026Updated last month
- low-latency realtime ASR based on FireRedASR☆62Jul 8, 2025Updated last year
- 最棒的的ASR后处理热词方案,基于音素编辑距离,实现热词替换。☆43Jun 10, 2026Updated last month
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆19,459Updated this week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Fun-ASR-Nano-2512的Docker版本☆16Jan 10, 2026Updated 6 months ago
- Official Python toolkit for the Qwen3-ASR API. Parallel high‑throughput calls, robust long‑audio transcription, multi‑sample‑rate support…☆981Feb 5, 2026Updated 5 months ago
- A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/…☆614Jun 2, 2026Updated last month
- Python runtime for WeTextProcessing (does not depend on Pynini)☆53Jun 11, 2026Updated last month
- 基于 FunASR SenseVoice 模型的实时语音识别服务,支持说话人识别、音频降噪、ASR 错误修正等高级功能。☆20Jul 10, 2026Updated 2 weeks ago
- Lightweight streaming Voice Activity Detection (VAD) tool with ONNX runtime☆24Mar 18, 2026Updated 4 months ago
- Pseudo Streaming SenseVoice with Hotwords☆466Jun 15, 2026Updated last month
- Local-first real-time meeting transcription with speaker diarization, switchable ASR engines, and optional OpenAI-compatible LLM summarie…☆85Updated this week
- A small and simple example showing how to run Qwen3-ASR with ONNX Runtime.☆33Apr 8, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- FastAPI to serve Qwen-ASR with streaming support. Tested. Benchmarked. Flash Attention 2. Fast & Stable.☆15Jun 24, 2026Updated last month
- An ASR API server for FunASR☆55Updated this week
- Dolphin is a multilingual, multitask ASR model jointly trained by DataoceanAI and Tsinghua University.☆776Jun 11, 2026Updated last month
- Implementation of Qwen3-ASR-0.6B in GGML☆101Updated this week
- X-ASR is a series of automatic speech recognition models based on the icefall framework, focusing on streaming ASR and low-latency deploy…☆145Jul 8, 2026Updated 2 weeks ago
- [ACL 2026 Main] Open-Ended Speaking Style Modeling via Fine-Grained and Multi-Granular Contrastive Language-Speech Pre-training☆104Apr 6, 2026Updated 3 months ago
- Utilizes ONNX Runtime to transcribe audio into text.☆85Jul 10, 2026Updated 2 weeks ago
- High-performance OCR microservice based on PaddleOCR-VL-0.9B (PaddleOCR-VL-1.5-0.9B) with MinerU-compatible API☆37Jan 30, 2026Updated 5 months ago
- GLM-ASR-Nano: A robust, open-source speech recognition model with 1.5B parameters☆836Mar 6, 2026Updated 4 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A lightweight demo of FunASR-Nano using ONNX runtime.☆83Feb 25, 2026Updated 5 months ago
- Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime…☆13,767Updated this week
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- Streaming Text to Speech Web UI☆22May 6, 2024Updated 2 years ago
- 将 Qwen3-ASR 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆196Apr 29, 2026Updated 2 months ago
- Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR be…☆1,940Feb 25, 2026Updated 5 months ago
- A Repository for Single- and Multi-modal Speaker Verification, Speaker Recognition and Speaker Diarization☆3,069Dec 8, 2025Updated 7 months ago