最极速的Qwen3-TTS推理方案。将 Qwen3-TTS 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。
☆168Jun 11, 2026Updated last month
Alternatives and similar repositories for Qwen3-TTS-GGUF
Users that are interested in Qwen3-TTS-GGUF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 将 Qwen3-ASR 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆203Apr 29, 2026Updated 3 months ago
- 用 onnx 和 gguf 格式混合运行 Fun-ASR-Nano 模型全流程☆153May 5, 2026Updated 2 months ago
- High-performance Qwen3-TTS implementation | Instruction-driven · Zero-shot voice cloning · Streaming · RTF 0.55☆68Jun 26, 2026Updated last month
- Real-time text-to-speech with Qwen3-TTS☆1,283Jul 17, 2026Updated 2 weeks ago
- Local AI text-to-speech with voice cloning and voice design, powered by GGML. C++17 port of OmniVoice (k2-fsa/OmniVoice). 646 languages, …☆150Jul 21, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆227Jul 18, 2026Updated 2 weeks ago
- SenseVoice-Small 导出为 ONNX,支持热词注入,在 CTC 的输空间中通过路径匹配,1ms 内实现热词替换☆28Jun 3, 2026Updated 2 months ago
- Fun-ASR is an end-to-end speech recognition large model launched by Tongyi Lab.☆109Jul 7, 2026Updated 3 weeks ago
- Easy fine-tuning for Qwen3-TTS: Fast voice cloning and high-quality multilingual speech synthesis.☆115May 29, 2026Updated 2 months ago
- About Samples code for Axera's PCIE Card for computer vision applications.☆20Feb 27, 2026Updated 5 months ago
- Lightweight streaming Voice Activity Detection (VAD) tool with ONNX runtime☆23Mar 18, 2026Updated 4 months ago
- 开箱即用的本地私有化部署语音服务,快速搭建Qwen3ASR/FunASR与Qwen3TTS/CosyVoice后端☆153Jul 6, 2026Updated 3 weeks ago
- Implementation of Qwen3-ASR-0.6B in GGML☆104Updated this week
- 实时城市路面积水面积检测(源码&教程)☆30Dec 6, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- AES Crypt for Windows is powerful file encryption software that integrates seamlessly with the Windows desktop and includes the AES Crypt…☆18Jun 1, 2026Updated 2 months ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- The cpp-based deployment of CosyVoice2☆20Sep 7, 2025Updated 10 months ago
- A lightweight demo of FunASR-Nano using ONNX runtime.☆82Feb 25, 2026Updated 5 months ago
- A curated list of models, benchmarks, tools and guides for audio editing☆35Updated this week
- Foley-Omni: a unified multimodal audio generation model for task-level synthesis and complete video soundtrack generation, producing spee…☆24Jun 5, 2026Updated last month
- Omni inference in C/C++☆230Updated this week
- Official repository of UniPASE, a SOTA USE model☆54Jul 21, 2026Updated last week
- A professional-grade interface for Qwen3-TTS, designed to unlock the model's full potential with fine-grained control and intuitive workf…☆286Mar 30, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A versatile and extendable web UI for FAL-AI, designed to seamlessly integrate with any FAL-AI API endpoint. This interface supports mode…☆18Sep 30, 2024Updated last year
- This project provides a production-ready, real-time inference server for LatentSync, enabling high-quality, low-latency 2D digital human …☆28Aug 16, 2025Updated 11 months ago
- Core Python library for MaterialSearch project.☆16Jul 20, 2026Updated 2 weeks ago
- A simple wrapper for hiroi-sora/PaddleOCR-json.☆17Oct 20, 2023Updated 2 years ago
- An N-gram punctuator for Chinese and English.☆20Oct 14, 2025Updated 9 months ago
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆12,778Mar 17, 2026Updated 4 months ago
- 使用vllm加速cosyvoice2的推理☆498Apr 26, 2025Updated last year
- Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music…☆3,290Jun 26, 2026Updated last month
- CosyVoice_DPO_NOTES: Supercharge Your Cosyvoice model with Cutting-Edge DPO Fine-Tuning!☆126Aug 8, 2025Updated 11 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆19,617Updated this week
- Covo-Audio is a 7B-parameter end-to-end large audio language model that directly processes continuous audio inputs and generates audio ou…☆176Mar 17, 2026Updated 4 months ago
- Speaker overlap-aware Neural Diarization☆12Feb 13, 2023Updated 3 years ago
- ☆14Aug 13, 2023Updated 2 years ago
- ☆17Apr 19, 2025Updated last year
- [ICLR2026] FlexiCodec: A Dynamic Neural Audio Codec for Low Frame Rates☆51Jul 1, 2026Updated last month
- Simple feed reader for Windows 11.☆41Aug 28, 2022Updated 3 years ago