用 onnx 和 gguf 格式混合运行 Fun-ASR-Nano 模型全流程
☆154May 5, 2026Updated 2 months ago
Alternatives and similar repositories for Fun-ASR-GGUF
Users that are interested in Fun-ASR-GGUF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 将 Qwen3-ASR 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆196Apr 29, 2026Updated 2 months ago
- A lightweight demo of FunASR-Nano using ONNX runtime.☆83Feb 25, 2026Updated 4 months ago
- Fun-ASR-Nano-2512官方发布的仓库内容有点多,部署起来坑也比较多,本项目提供一个简化的部署方案。☆150Dec 26, 2025Updated 6 months ago
- Utilizes ONNX Runtime to transcribe audio into text.☆84Jul 10, 2026Updated last week
- 最极速的Qwen3-TTS推理方案。将 Qwen3-TTS 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆162Jun 11, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp…☆1,412Updated this week
- SenseVoice-Small 导出为 ONNX,支持热词注入,在 CTC 的输空间中通过路径匹配,1ms 内实现热词替换☆28Jun 3, 2026Updated last month
- Fun-ASR is an end-to-end speech recognition large model launched by Tongyi Lab.☆107Jul 7, 2026Updated 2 weeks ago
- ☆49Mar 18, 2026Updated 4 months ago
- Uses ONNX Runtime for character role speaker identification.☆17Dec 28, 2025Updated 6 months ago
- State-of-the-art continious audio tokenization☆40Mar 9, 2026Updated 4 months ago
- Source code for "BLOOM-Net: Blockwise Optimization for Masking Networks Toward Scalable and Efficient Speech Enhancement"☆14Feb 13, 2022Updated 4 years ago
- Transcribe subtitles and translate them offline with ease.☆47Jun 15, 2026Updated last month
- X-ASR is a series of automatic speech recognition models based on the icefall framework, focusing on streaming ASR and low-latency deploy…☆142Jul 8, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music…☆3,191Jun 26, 2026Updated 3 weeks ago
- Demonstrate Yolov9 model with Qualcomm Hexagon NPU and DirectML☆12Nov 27, 2024Updated last year
- The inference and trainging code for WordVoice.☆48Updated this week
- The dataset construction pipeline for WordVoice-5A☆15Updated this week
- Lightweight streaming Voice Activity Detection (VAD) tool with ONNX runtime☆22Mar 18, 2026Updated 4 months ago
- ☆17Mar 30, 2023Updated 3 years ago
- ☆19Jan 6, 2025Updated last year
- A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/…☆607Jun 2, 2026Updated last month
- 中文逆文本正则化 (Chinese ITN, Chinese Inverse Text Normalization) ,即将文本中的中文数字转为阿拉伯数字。☆32Jun 10, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- FlashCosyVoice: A lightweight vLLM implementation built from scratch for CosyVoice.☆250Feb 25, 2026Updated 4 months ago
- Utilizes ONNX Runtime for TTS model.☆65Jul 13, 2026Updated last week
- The baselines of ARC-Challenge-Interspeech2026☆60Dec 1, 2025Updated 7 months ago
- Generative_Annotation_NEC: A novel NEC method that utilizes speech sound features to retrieve candidate entities and a generative method …☆17Dec 2, 2025Updated 7 months ago
- ☆47Apr 27, 2026Updated 2 months ago
- Utilizes ONNX Runtime for speech activity detection.☆46Jun 25, 2026Updated 3 weeks ago
- Standalone C++ inference project for VoxCPM models built on top of ggml.☆84Jul 14, 2026Updated last week
- low-latency realtime ASR based on FireRedASR☆62Jul 8, 2025Updated last year
- ☆16Nov 4, 2025Updated 8 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆34Sep 25, 2025Updated 9 months ago
- Port of Funasr's Sense-voice model in C/C++☆568Dec 19, 2025Updated 7 months ago
- X-VC: Zero-shot Streaming Voice Conversion in Codec Space☆69May 6, 2026Updated 2 months ago
- ☆24Jun 10, 2026Updated last month
- The implementation of G2Net, the extension of GaGNet and is in submission to T-ASLP☆19Apr 27, 2022Updated 4 years ago
- ☆35Sep 6, 2025Updated 10 months ago
- Official implementation: "AudioSet-R: A Refined AudioSet with Multi-Stage LLM Label Reannotation"☆19Oct 9, 2025Updated 9 months ago