An open-source model family for long-form speech, dialogue synthesis, voice design, sound effects, and real-time streaming TTS
☆4,147Sep 6, 2026Updated 3 weeks ago
Alternatives and similar repositories for MOSS-TTS
Users that are interested in MOSS-TTS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A 1.6B causal Transformer audio tokenizer with streaming, variable bitrates, and semantic alignment across speech, sound, and music☆258Jun 16, 2026Updated 3 months ago
- A 100M-parameter multilingual TTS model for real-time CPU inference, voice cloning, and 48 kHz stereo generation☆4,428Sep 6, 2026Updated 3 weeks ago
- VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning☆38,133Sep 2, 2026Updated 3 weeks ago
- A multilingual model for long-form, multi-speaker dialogue synthesis with flexible speaker control and zero-shot voice cloning☆1,400Sep 6, 2026Updated 3 weeks ago
- High-Quality Voice Cloning TTS for 600+ Languages☆14,028Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆1,349Updated this week
- An open-source model for understanding speech, environmental sounds, and music through captioning, question answering, and reasoning☆675Sep 6, 2026Updated 3 weeks ago
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆13,591Mar 17, 2026Updated 6 months ago
- An open-weight 11B model series for long-form and real-time video understanding☆743Updated this week
- Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.☆13,789Sep 9, 2026Updated 3 weeks ago
- Ming-omni-tts: Simple and Efficient Unified Generation of Speech, Music, and Sound with Precise Control☆265Feb 26, 2026Updated 7 months ago
- A foundation model that generates synchronized video and audio in a single model☆1,119Updated this week
- SOTA Open Source TTS☆32,886Sep 16, 2026Updated 2 weeks ago
- Open-Source Frontier Voice AI☆54,538Sep 3, 2026Updated 3 weeks ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Official inference code for SoulX-Singer: Towards High-Quality Zero-Shot Singing Voice Synthesis☆977May 29, 2026Updated 4 months ago
- ☆579Apr 3, 2026Updated 5 months ago
- GLM-TTS: Controllable & Emotion-Expressive Zero-shot TTS with Multi-Reward Reinforcement Learning☆1,066Apr 10, 2026Updated 5 months ago
- A TTS that fits in your CPU (and pocket)☆9,693Updated this week
- The open-source AI voice studio. Clone, dictate, create.☆55,971Aug 9, 2026Updated last month
- ☆218Jun 2, 2026Updated 3 months ago
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆23,775May 25, 2026Updated 4 months ago
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching☆1,075Dec 2, 2025Updated 9 months ago
- "ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"☆12,516Sep 20, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics…☆980Apr 9, 2026Updated 5 months ago
- WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling☆212Jun 6, 2026Updated 3 months ago
- An end-to-end speech-to-speech language model that generates spoken responses without text guidance☆139Feb 13, 2026Updated 7 months ago
- SoTA open-source TTS☆26,612Jul 21, 2026Updated 2 months ago
- An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System☆24,226Updated this week
- super expressive prompting model based on ltx2.3☆496May 23, 2026Updated 4 months ago
- [ACL 2026 Main] Training, inference, and testing of the SAC speech codec model.☆111Nov 1, 2025Updated 10 months ago
- A 0.9B model for long-form transcription in 50+ languages with speaker diarization, timestamps, and acoustic event awareness☆2,092Sep 15, 2026Updated 2 weeks ago
- On-device TTS model by Neuphonic☆6,297Jul 30, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation☆457Nov 27, 2025Updated 10 months ago
- Zonos2 is a leading open-weight text-to-speech MoE.☆314Updated this week
- Open source voice AI platform. Self-hosted alternative to Vapi and Retell. On Prem, BYOK across Speech to Speech or LLM/STT/TTS, with a …☆5,773Updated this week
- MiMo-Audio: Audio Language Models are Few-Shot Learners☆1,083Jun 17, 2026Updated 3 months ago
- Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model.☆9,543Aug 26, 2026Updated last month
- 🌋LavaSR: Fast Speech restoration and enhancement☆601Jun 19, 2026Updated 3 months ago
- ☆303Jul 22, 2025Updated last year