将 Qwen3-ASR 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。
☆236Apr 29, 2026Updated 4 months ago
Alternatives and similar repositories for Qwen3-ASR-GGUF
Users that are interested in Qwen3-ASR-GGUF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 用 onnx 和 gguf 格式混合运行 Fun-ASR-Nano 模型全流程☆161May 5, 2026Updated 4 months ago
- 最极速的Qwen3-TTS推理方案。将 Qwen3-TTS 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆190Updated this week
- Qwen3-ASR speech-to-text for llama.cpp — patch, GGUF models, and benchmarks☆18Feb 2, 2026Updated 7 months ago
- Implementation of Qwen3-ASR-0.6B in GGML☆115Jul 28, 2026Updated last month
- A lightweight demo of FunASR-Nano using ONNX runtime.☆87Feb 25, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Pure-Rust inference engine for Qwen3-ASR speech recognition models (0.6B & 1.7B) using candle with Metal/CUDA acceleration☆27Mar 17, 2026Updated 6 months ago
- Fun-ASR is an end-to-end speech recognition large model launched by Tongyi Lab.☆111Jul 7, 2026Updated 2 months ago
- SenseVoice-Small 导出为 ONNX,支持热词注入,在 CTC 的输空间中通过路径匹配,1ms 内实现热词替换☆32Jun 3, 2026Updated 3 months ago
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆17Jun 27, 2026Updated 2 months ago
- Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp…☆1,551Sep 10, 2026Updated last week
- AI Speech Solutions for Tasks such as ASR, Vocal Extraction, Accompaniment Extraction, Audio Denoising, and Enhancement, Support models s…☆86Jun 16, 2026Updated 3 months ago
- ☆241Jul 18, 2026Updated 2 months ago
- Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music…☆3,582Jun 26, 2026Updated 2 months ago
- ☆19Mar 18, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- State-of-the-art continious audio tokenization☆42Mar 9, 2026Updated 6 months ago
- X-ASR is a series of automatic speech recognition models based on the icefall framework, focusing on streaming ASR and low-latency deploy…☆185Jul 29, 2026Updated last month
- C inference for Qwen3-ASR 0.6b and 1.7b transcriptions models☆615Feb 17, 2026Updated 7 months ago
- 中文逆文本正则化 (Chinese ITN, Chinese Inverse Text Normalization) ,即将文本中的中文数字转为阿拉伯数字。☆36Sep 7, 2026Updated 2 weeks ago
- A small and simple example showing how to run Qwen3-ASR with ONNX Runtime.☆40Apr 8, 2026Updated 5 months ago
- Utilizes ONNX Runtime to transcribe audio into text.☆89Aug 26, 2026Updated 3 weeks ago
- 基于FunASR官方Demo修改的WS服务端,配合FastAPI提供HTTP服务,可以在浏览器中进行实时ASR测试☆57Aug 4, 2025Updated last year
- [ICASSP 2026] Official code for "Measuring Prosody Diversity in Zero-Shot TTS: A New Metric, Benchmark, and Exploration"☆18Apr 16, 2026Updated 5 months ago
- Official Python toolkit for the Qwen3-ASR API. Parallel high‑throughput calls, robust long‑audio transcription, multi‑sample‑rate support…☆1,009Feb 5, 2026Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Lightweight streaming Voice Activity Detection (VAD) tool with ONNX runtime☆24Mar 18, 2026Updated 6 months ago
- A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/…☆690Jun 2, 2026Updated 3 months ago
- Dolphin is a multilingual, multitask ASR model jointly trained by DataoceanAI and Tsinghua University.☆791Jun 11, 2026Updated 3 months ago
- Port of Funasr's Sense-voice model in C/C++☆575Dec 19, 2025Updated 9 months ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- Whisper-based voice activity detection toolkit covering encoder-only, DETR, and WhisperSeg refinements with Lightning training, ONNX expo…☆24Nov 24, 2025Updated 9 months ago
- C++ implementation of "Mobile Vision Transformer-based Visual Object Tracking" (BMVC2023) and "Separable Self and Mixed Attention Transf…☆13Apr 23, 2024Updated 2 years ago
- 一个模块化,全过程可离线,低占用率的对话机器人/智能音箱☆162Mar 25, 2026Updated 5 months ago
- PC 端语音输入工具,离线识别,高准确率、低延迟,支持热词、LLM润色。按住CapsLock或鼠标侧键X2说话,松开自动上屏。☆6,848Sep 14, 2026Updated last week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal for…☆660Updated this week
- ☆18Apr 25, 2026Updated 4 months ago
- ☆53Mar 18, 2026Updated 6 months ago
- sherpa-onnx Go package for Windows☆14Sep 11, 2026Updated last week
- Utilizes ONNX Runtime for TTS model.☆70Sep 4, 2026Updated 2 weeks ago
- Multi-provider ASR web studio (Qwen, Doubao, Gemini, NIM, OpenAI-compatible & more) with recording, batch queue, PWA, local cache, and be…☆276Jul 16, 2026Updated 2 months ago
- aha model inference library, now supports Qwen(2.5VL/3/3VL/3.5/ASR/3Embedding/3Reranker), MiniCPM(4/5), VoxCPM(0.5B/1.5/2), DeepSeek-OCR/…☆394Jun 7, 2026Updated 3 months ago