最极速的Qwen3-TTS推理方案。将 Qwen3-TTS 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。
☆183Sep 5, 2026Updated last week
Alternatives and similar repositories for Qwen3-TTS-GGUF
Users that are interested in Qwen3-TTS-GGUF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 将 Qwen3-ASR 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆230Apr 29, 2026Updated 4 months ago
- 用 onnx 和 gguf 格式混合运行 Fun-ASR-Nano 模型全流程☆161May 5, 2026Updated 4 months ago
- ☆18May 13, 2026Updated 3 months ago
- High-performance Qwen3-TTS implementation | Instruction-driven · Zero-shot voice cloning · Streaming · RTF 0.55☆67Jun 26, 2026Updated 2 months ago
- Real-time text-to-speech with Qwen3-TTS☆1,347Aug 25, 2026Updated 2 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Local AI text-to-speech with voice cloning and voice design, powered by GGML. C++17 port of OmniVoice (k2-fsa/OmniVoice). 646 languages, …☆172Updated this week
- ☆240Jul 18, 2026Updated last month
- SenseVoice-Small 导出为 ONNX,支持热词注入,在 CTC 的输空间中通过路径匹配,1ms 内实现热词替换☆30Jun 3, 2026Updated 3 months ago
- FlowMirror-HydraVox — A natively accelerated multi-head autoregressive TTS system derived from CosyVoice 3.0. It predicts multiple tokens…☆49Feb 17, 2026Updated 6 months ago
- Easy fine-tuning for Qwen3-TTS: Fast voice cloning and high-quality multilingual speech synthesis.☆122May 29, 2026Updated 3 months ago
- About Samples code for Axera's PCIE Card for computer vision applications.☆20Aug 10, 2026Updated last month
- 开箱即用的本地私有化部署语音服务,快速搭建Qwen3ASR/FunASR与Qwen3TTS/CosyVoice后端☆163Jul 6, 2026Updated 2 months ago
- Implementation of Qwen3-ASR-0.6B in GGML☆111Jul 28, 2026Updated last month
- AES Crypt for Windows is powerful file encryption software that integrates seamlessly with the Windows desktop and includes the AES Crypt…☆18Jun 1, 2026Updated 3 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 🚀 A centralized launcher for managing and running BAT and PS1 scripts.☆18Jun 20, 2026Updated 2 months ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- A lightweight demo of FunASR-Nano using ONNX runtime.☆87Feb 25, 2026Updated 6 months ago
- C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal for…☆626Updated this week
- A curated list of models, benchmarks, tools and guides for audio editing☆46Updated this week
- Foley-Omni: a unified multimodal audio generation model for task-level synthesis and complete video soundtrack generation, producing spee…☆27Jun 5, 2026Updated 3 months ago
- Omni inference in C/C++☆260Updated this week
- A professional-grade interface for Qwen3-TTS, designed to unlock the model's full potential with fine-grained control and intuitive workf…☆294Mar 30, 2026Updated 5 months ago
- A versatile and extendable web UI for FAL-AI, designed to seamlessly integrate with any FAL-AI API endpoint. This interface supports mode…☆18Sep 30, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official repository of UniPASE, a SOTA USE model☆71Aug 14, 2026Updated 3 weeks ago
- This project provides a production-ready, real-time inference server for LatentSync, enabling high-quality, low-latency 2D digital human …☆30Aug 16, 2025Updated last year
- api document for www.xt.com , www.xt.pub etc☆10Jun 17, 2022Updated 4 years ago
- An N-gram punctuator for Chinese and English.☆20Oct 14, 2025Updated 10 months ago
- Local AI text-to-speech with voice cloning and voice design, powered by GGML. C++17 port of Qwen3-TTS (QwenLM/Qwen3-TTS). 10 languages, 2…☆162Updated this week
- Official code for MUSE: Flexible Voiceprint Receptive Fields and Multi-Path Fusion Enhanced Taylor Transformer for U-Net-based Speech Enh…☆58Mar 5, 2025Updated last year
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆13,362Mar 17, 2026Updated 5 months ago
- Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music…☆3,517Jun 26, 2026Updated 2 months ago
- 自分用のカスタムノード☆15Jun 6, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆20,285Updated this week
- Speaker overlap-aware Neural Diarization☆12Feb 13, 2023Updated 3 years ago
- A GUI proxy client in GTK+3. Based on Clash.☆14Feb 12, 2019Updated 7 years ago
- [ICLR2026] FlexiCodec: A Dynamic Neural Audio Codec for Low Frame Rates☆52Aug 20, 2026Updated 3 weeks ago
- make your charactor Dancing as Native style☆40Dec 16, 2023Updated 2 years ago
- Simple feed reader for Windows 11.☆39Aug 28, 2022Updated 4 years ago
- 几年前读Live555的源码的时候做的注释☆11Jan 17, 2019Updated 7 years ago