将 Qwen3-ASR 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。
☆226Apr 29, 2026Updated 4 months ago
Alternatives and similar repositories for Qwen3-ASR-GGUF
Users that are interested in Qwen3-ASR-GGUF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 用 onnx 和 gguf 格式混合运行 Fun-ASR-Nano 模型全流程☆160May 5, 2026Updated 3 months ago
- 最极速的Qwen3-TTS推理方案。将 Qwen3-TTS 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆182Updated this week
- Qwen3-ASR speech-to-text for llama.cpp — patch, GGUF models, and benchmarks☆17Feb 2, 2026Updated 7 months ago
- Implementation of Qwen3-ASR-0.6B in GGML☆110Jul 28, 2026Updated last month
- A lightweight demo of FunASR-Nano using ONNX runtime.☆85Feb 25, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Pure-Rust inference engine for Qwen3-ASR speech recognition models (0.6B & 1.7B) using candle with Metal/CUDA acceleration☆25Mar 17, 2026Updated 5 months ago
- Fun-ASR is an end-to-end speech recognition large model launched by Tongyi Lab.☆110Jul 7, 2026Updated last month
- SenseVoice-Small 导出为 ONNX,支持热词注入,在 CTC 的输空间中通过路径匹配,1ms 内实现热词替换☆30Jun 3, 2026Updated 3 months ago
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆16Jun 27, 2026Updated 2 months ago
- Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp…☆1,512Updated this week
- AI Speech Solutions for Tasks such as ASR, Vocal Extraction, Accompaniment Extraction, Audio Denoising, and Enhancement, Support models s…☆85Jun 16, 2026Updated 2 months ago
- ☆235Jul 18, 2026Updated last month
- ☆19Mar 18, 2026Updated 5 months ago
- Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music…☆3,459Jun 26, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- State-of-the-art continious audio tokenization☆42Mar 9, 2026Updated 5 months ago
- X-ASR is a series of automatic speech recognition models based on the icefall framework, focusing on streaming ASR and low-latency deploy…☆173Jul 29, 2026Updated last month
- C inference for Qwen3-ASR 0.6b and 1.7b transcriptions models☆602Feb 17, 2026Updated 6 months ago
- 中文逆文本正则化 (Chinese ITN, Chinese Inverse Text Normalization) ,即将文本中的中文数字转为阿拉伯数字。☆33Aug 12, 2026Updated 3 weeks ago
- A small and simple example showing how to run Qwen3-ASR with ONNX Runtime.☆36Apr 8, 2026Updated 4 months ago
- Utilizes ONNX Runtime to transcribe audio into text.☆87Aug 26, 2026Updated last week
- [ICASSP 2026] Official code for "Measuring Prosody Diversity in Zero-Shot TTS: A New Metric, Benchmark, and Exploration"☆17Apr 16, 2026Updated 4 months ago
- 基于FunASR官方Demo修改的WS服务端,配合FastAPI提供HTTP服务,可以在浏览器中进行实时ASR测试☆57Aug 4, 2025Updated last year
- Official Python toolkit for the Qwen3-ASR API. Parallel high‑throughput calls, robust long‑audio transcription, multi‑sample‑rate support…☆1,002Feb 5, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Lightweight streaming Voice Activity Detection (VAD) tool with ONNX runtime☆24Mar 18, 2026Updated 5 months ago
- A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/…☆665Jun 2, 2026Updated 3 months ago
- Dolphin is a multilingual, multitask ASR model jointly trained by DataoceanAI and Tsinghua University.☆786Jun 11, 2026Updated 2 months ago
- Port of Funasr's Sense-voice model in C/C++☆573Dec 19, 2025Updated 8 months ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- Whisper-based voice activity detection toolkit covering encoder-only, DETR, and WhisperSeg refinements with Lightning training, ONNX expo…☆22Nov 24, 2025Updated 9 months ago
- C++ implementation of "Mobile Vision Transformer-based Visual Object Tracking" (BMVC2023) and "Separable Self and Mixed Attention Transf…☆13Apr 23, 2024Updated 2 years ago
- 一个模块化,全过程可离线,低占用率的对话机器人/智能音箱☆162Mar 25, 2026Updated 5 months ago
- PC 端语音输入工具,离线识别,高准确率、低延迟,支持热 词、LLM润色。按住CapsLock或鼠标侧键X2说话,松开自动上屏。☆6,736Aug 26, 2026Updated last week
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal for…☆600Updated this week
- ☆50Mar 18, 2026Updated 5 months ago
- sherpa-onnx Go package for Windows☆14Updated this week
- Utilizes ONNX Runtime for TTS model.☆70Aug 26, 2026Updated last week
- Multi-provider ASR web studio (Qwen, Doubao, Gemini, NIM, OpenAI-compatible & more) with recording, batch queue, PWA, local cache, and be…☆274Jul 16, 2026Updated last month
- aha model inference library, now supports Qwen(2.5VL/3/3VL/3.5/ASR/3Embedding/3Reranker), MiniCPM(4/5), VoxCPM(0.5B/1.5/2), DeepSeek-OCR/…☆391Jun 7, 2026Updated 2 months ago
- GLM-ASR-Nano: A robust, open-source speech recognition model with 1.5B parameters☆853Mar 6, 2026Updated 5 months ago