用 onnx 和 gguf 格式混合运行 Fun-ASR-Nano 模型全流程
☆161May 5, 2026Updated 4 months ago
Alternatives and similar repositories for Fun-ASR-GGUF
Users that are interested in Fun-ASR-GGUF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 将 Qwen3-ASR 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆236Apr 29, 2026Updated 4 months ago
- A lightweight demo of FunASR-Nano using ONNX runtime.☆87Feb 25, 2026Updated 6 months ago
- Fun-ASR-Nano-2512官方发布的仓库内容有点多,部署起来坑也比较多,本项目提供一个简化的部署方案。☆151Dec 26, 2025Updated 8 months ago
- Utilizes ONNX Runtime to transcribe audio into text.☆89Aug 26, 2026Updated 3 weeks ago
- 最极速的Qwen3-TTS推理方案。将 Qwen3-TTS 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆190Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp…☆1,551Sep 10, 2026Updated last week
- SenseVoice-Small 导出为 ONNX,支持热词注入,在 CTC 的输空间中通过路径匹配,1ms 内实现热词替换☆32Jun 3, 2026Updated 3 months ago
- Fun-ASR is an end-to-end speech recognition large model launched by Tongyi Lab.☆111Jul 7, 2026Updated 2 months ago
- ☆53Mar 18, 2026Updated 6 months ago
- Uses ONNX Runtime for character role speaker identification.☆17Dec 28, 2025Updated 8 months ago
- State-of-the-art continious audio tokenization☆42Mar 9, 2026Updated 6 months ago
- Source code for "BLOOM-Net: Blockwise Optimization for Masking Networks Toward Scalable and Efficient Speech Enhancement"☆14Feb 13, 2022Updated 4 years ago
- Utilizes ONNX Runtime for speech activity detection.☆50Aug 17, 2026Updated last month
- X-ASR is a series of automatic speech recognition models based on the icefall framework, focusing on streaming ASR and low-latency deploy…☆185Jul 29, 2026Updated last month
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music…☆3,582Jun 26, 2026Updated 2 months ago
- Demonstrate Yolov9 model with Qualcomm Hexagon NPU and DirectML☆12Nov 27, 2024Updated last year
- Port of Funasr's Sense-voice model in C/C++☆575Dec 19, 2025Updated 9 months ago
- The dataset construction pipeline for WordVoice-5A☆19Jul 17, 2026Updated 2 months ago
- ☆17Mar 30, 2023Updated 3 years ago
- Lightweight streaming Voice Activity Detection (VAD) tool with ONNX runtime☆24Mar 18, 2026Updated 6 months ago
- ☆19Jan 6, 2025Updated last year
- 中文逆文本正则化 (Chinese ITN, Chinese Inverse Text Normalization) ,即将文本中的中文数字转为阿拉伯数字。☆36Sep 7, 2026Updated 2 weeks ago
- A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/…☆690Jun 2, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- FlashCosyVoice: A lightweight vLLM implementation built from scratch for CosyVoice.☆253Feb 25, 2026Updated 6 months ago
- Utilizes ONNX Runtime for TTS model.☆70Sep 4, 2026Updated 2 weeks ago
- Transcribe subtitles and translate them offline with ease.☆47Jun 15, 2026Updated 3 months ago
- The baselines of ARC-Challenge-Interspeech2026☆62Dec 1, 2025Updated 9 months ago
- Generative_Annotation_NEC: A novel NEC method that utilizes speech sound features to retrieve candidate entities and a generative method …☆18Dec 2, 2025Updated 9 months ago
- ☆52Apr 27, 2026Updated 4 months ago
- Standalone C++ inference project for VoxCPM models built on top of ggml.☆94Jul 14, 2026Updated 2 months ago
- low-latency realtime ASR based on FireRedASR☆61Jul 8, 2025Updated last year
- The inference and trainging code for WordVoice.☆72Aug 21, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆35Sep 25, 2025Updated 11 months ago
- ☆17Nov 4, 2025Updated 10 months ago
- X-VC: Zero-shot Streaming Voice Conversion in Codec Space☆80May 6, 2026Updated 4 months ago
- The implementation of G2Net, the extension of GaGNet and is in submission to T-ASLP☆19Apr 27, 2022Updated 4 years ago
- ☆28Jun 10, 2026Updated 3 months ago
- ☆35Sep 6, 2025Updated last year
- Official implementation: "AudioSet-R: A Refined AudioSet with Multi-Stage LLM Label Reannotation"☆20Oct 9, 2025Updated 11 months ago