用 onnx 和 gguf 格式混合运行 Fun-ASR-Nano 模型全流程
☆154May 5, 2026Updated 3 months ago
Alternatives and similar repositories for Fun-ASR-GGUF
Users that are interested in Fun-ASR-GGUF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 将 Qwen3-ASR 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆207Apr 29, 2026Updated 3 months ago
- A lightweight demo of FunASR-Nano using ONNX runtime.☆82Feb 25, 2026Updated 5 months ago
- Fun-ASR-Nano-2512官方发布的仓库内容有点多,部署起来坑也比较多,本项目提供一个简化的部署方案。☆150Dec 26, 2025Updated 7 months ago
- Utilizes ONNX Runtime to transcribe audio into text.☆85Aug 1, 2026Updated last week
- 最极速的Qwen3-TTS推理方案。将 Qwen3-TTS 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆171Jun 11, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp…☆1,464Jul 24, 2026Updated 2 weeks ago
- SenseVoice-Small 导出为 ONNX,支持热词注入,在 CTC 的输空间中通过路径匹配,1ms 内实现热词替换☆29Jun 3, 2026Updated 2 months ago
- Fun-ASR is an end-to-end speech recognition large model launched by Tongyi Lab.☆109Jul 7, 2026Updated last month
- ☆49Mar 18, 2026Updated 4 months ago
- Uses ONNX Runtime for character role speaker identification.☆17Dec 28, 2025Updated 7 months ago
- State-of-the-art continious audio tokenization☆41Mar 9, 2026Updated 5 months ago
- Source code for "BLOOM-Net: Blockwise Optimization for Masking Networks Toward Scalable and Efficient Speech Enhancement"☆14Feb 13, 2022Updated 4 years ago
- Utilizes ONNX Runtime for speech activity detection.☆47Aug 4, 2026Updated last week
- X-ASR is a series of automatic speech recognition models based on the icefall framework, focusing on streaming ASR and low-latency deploy…☆159Jul 29, 2026Updated last week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music…☆3,337Jun 26, 2026Updated last month
- Demonstrate Yolov9 model with Qualcomm Hexagon NPU and DirectML☆12Nov 27, 2024Updated last year
- Port of Funasr's Sense-voice model in C/C++☆571Dec 19, 2025Updated 7 months ago
- The dataset construction pipeline for WordVoice-5A☆19Jul 17, 2026Updated 3 weeks ago
- ☆17Mar 30, 2023Updated 3 years ago
- Lightweight streaming Voice Activity Detection (VAD) tool with ONNX runtime☆23Mar 18, 2026Updated 4 months ago
- ☆19Jan 6, 2025Updated last year
- A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/…☆631Jun 2, 2026Updated 2 months ago
- 中文逆文本正则化 (Chinese ITN, Chinese Inverse Text Normalization) ,即将文本中的中文数字转为阿拉伯数字。☆33Jun 10, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- FlashCosyVoice: A lightweight vLLM implementation built from scratch for CosyVoice.☆250Feb 25, 2026Updated 5 months ago
- Utilizes ONNX Runtime for TTS model.☆65Jul 30, 2026Updated last week
- Transcribe subtitles and translate them offline with ease.☆47Jun 15, 2026Updated last month
- The baselines of ARC-Challenge-Interspeech2026☆60Dec 1, 2025Updated 8 months ago
- Generative_Annotation_NEC: A novel NEC method that utilizes speech sound features to retrieve candidate entities and a generative method …☆17Dec 2, 2025Updated 8 months ago
- ☆48Apr 27, 2026Updated 3 months ago
- Standalone C++ inference project for VoxCPM models built on top of ggml.☆89Jul 14, 2026Updated 3 weeks ago
- low-latency realtime ASR based on FireRedASR☆61Jul 8, 2025Updated last year
- The inference and trainging code for WordVoice.☆68Jul 30, 2026Updated last week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆35Sep 25, 2025Updated 10 months ago
- ☆16Nov 4, 2025Updated 9 months ago
- X-VC: Zero-shot Streaming Voice Conversion in Codec Space☆73May 6, 2026Updated 3 months ago
- ☆24Jun 10, 2026Updated 2 months ago
- The implementation of G2Net, the extension of GaGNet and is in submission to T-ASLP☆19Apr 27, 2022Updated 4 years ago
- ☆35Sep 6, 2025Updated 11 months ago
- Official implementation: "AudioSet-R: A Refined AudioSet with Multi-Stage LLM Label Reannotation"☆20Oct 9, 2025Updated 10 months ago