Real-time text-to-speech with Qwen3-TTS
☆1,315Jul 17, 2026Updated last month
Alternatives and similar repositories for faster-qwen3-tts
Users that are interested in faster-qwen3-tts are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Fork with streaming inference support + ~6× faster inference☆259Jun 3, 2026Updated 2 months ago
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆13,007Mar 17, 2026Updated 5 months ago
- Qwen3-TTS with nano vLLM-style optimizations for fast text-to-speech generation. Achieved 3x faster☆137Mar 3, 2026Updated 5 months ago
- Real-time streaming TTS for Qwen3-TTS with two-phase latency, Hann crossfade, and torch.compile + CUDA graphs optimizations☆96Feb 21, 2026Updated 5 months ago
- High-Quality Voice Cloning TTS for 600+ Languages☆9,220Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 最极速的Qwen3-TTS推理方案。将 Qwen3-TTS 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆177Jun 11, 2026Updated 2 months ago
- Pure-PyTorch Parakeet TDT inference☆52Mar 10, 2026Updated 5 months ago
- MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fi…☆4,005Jul 26, 2026Updated 3 weeks ago
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching☆1,039Dec 2, 2025Updated 8 months ago
- Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music…☆3,383Jun 26, 2026Updated last month
- A SOTA Industrial-Grade Voice Activity Detection & Audio Event Detection, supporting 100+ languages, outperforming Silero-VAD, TEN-VAD, F…☆505May 6, 2026Updated 3 months ago
- Soprano: Instant, Ultra-Realistic Text-to-Speech☆1,481Jan 15, 2026Updated 7 months ago
- Ming-omni-tts: Simple and Efficient Unified Generation of Speech, Music, and Sound with Precise Control☆265Feb 26, 2026Updated 5 months ago
- Open Source Speech Language Model☆1,008May 11, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A high quality and fast TTS repository☆520Dec 22, 2025Updated 7 months ago
- Triton kernel fusion for Qwen3-TTS 1.7B inference acceleration — RMSNorm, SwiGLU, M-RoPE, Norm+Residual☆104Jun 29, 2026Updated last month
- A high-quality rapid TTS voice cloning model that reaches speeds of 150x realtime.☆5,226Jun 5, 2026Updated 2 months ago
- A TTS that fits in your CPU (and pocket)☆8,705Updated this week
- Soprano-Factory: Train your own 2000x realtime text-to-speech model☆254Jan 13, 2026Updated 7 months ago
- Build local voice agents with open-source models☆12,607Updated this week
- ☆1,202Updated this week
- Rust implementation of Qwen3-ASR automatic speech recognition☆250Mar 28, 2026Updated 4 months ago
- A simple implementation for improving CosyVoice2 by GRPO method☆39May 5, 2026Updated 3 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- LEMAS‑TTS is a multilingual zero‑shot text‑to‑speech system, supporting 10 languages: Chinese English Spanish Russian French German Ital…☆102Mar 31, 2026Updated 4 months ago
- GLM-TTS: Controllable & Emotion-Expressive Zero-shot TTS with Multi-Reward Reinforcement Learning☆1,051Apr 10, 2026Updated 4 months ago
- ☆565Apr 3, 2026Updated 4 months ago
- Fast audio super resolution from 16khz to 48khz.☆217Jan 3, 2026Updated 7 months ago
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆222Jul 10, 2026Updated last month
- Easy fine-tuning for Qwen3-TTS: Fast voice cloning and high-quality multilingual speech synthesis.☆120May 29, 2026Updated 2 months ago
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆22,809May 25, 2026Updated 2 months ago
- Zonos2 is a leading open-weight text-to-speech MoE.☆298Jul 6, 2026Updated last month
- ☆303Jul 22, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A framework for efficient model inference with omni-modality models☆6,150Updated this week
- Towards Human-Sounding Speech☆6,297Dec 5, 2025Updated 8 months ago
- SOTA Open Source TTS☆32,249Aug 3, 2026Updated 2 weeks ago
- ☆192Aug 25, 2025Updated 11 months ago
- VyvoTTS: LLM-Based Text-to-Speech Training Framework☆261Aug 9, 2026Updated last week
- On-device TTS model by Neuphonic☆6,239Jul 30, 2026Updated 2 weeks ago
- Interface for OuteTTS models.☆1,435Mar 23, 2026Updated 4 months ago