MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and complex real‑world scenarios, covering stable long‑form speech, multi‑speaker dialogue, voice/character design, environmental sound effects, and real‑time streaming TTS.
☆4,005Jul 26, 2026Updated 3 weeks ago
Alternatives and similar repositories for MOSS-TTS
Users that are interested in MOSS-TTS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MOSS-Audio-Tokenizer is a Causal Transformer-based audio tokenizer built on the CAT architecture. Trained on 3M hours of diverse audio, i…☆249Jun 16, 2026Updated 2 months ago
- MOSS-TTS-Nano is an open-source multilingual tiny speech generation model from MOSI.AI and the OpenMOSS team. With only 0.1B parameters, …☆4,191Jul 26, 2026Updated 3 weeks ago
- VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning☆35,816Aug 12, 2026Updated last week
- MOSS-TTSD is a spoken dialogue generation model designed for expressive multi-speaker synthesis. It features long-context modeling, flex…☆1,385Jul 26, 2026Updated 3 weeks ago
- High-Quality Voice Cloning TTS for 600+ Languages☆9,220Updated this week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆1,202Updated this week
- MOSS-Audio is an open-source foundation model for unified audio understanding, enabling speech, sound, music, captioning, QA, and reasoni…☆639Jun 2, 2026Updated 2 months ago
- MOSS-VL is the core multimodal model series within the OpenMOSS ecosystem, dedicated to visual understanding.☆435Updated this week
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆13,007Mar 17, 2026Updated 5 months ago
- Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.☆13,691Jul 24, 2026Updated 3 weeks ago
- MOVA: Towards Scalable and Synchronized Video–Audio Generation☆1,098Updated this week
- SOTA Open Source TTS☆32,249Aug 3, 2026Updated 2 weeks ago
- Ming-omni-tts: Simple and Efficient Unified Generation of Speech, Music, and Sound with Precise Control☆265Feb 26, 2026Updated 5 months ago
- Open-Source Frontier Voice AI☆52,873Jul 24, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official inference code for SoulX-Singer: Towards High-Quality Zero-Shot Singing Voice Synthesis☆922May 29, 2026Updated 2 months ago
- ☆565Apr 3, 2026Updated 4 months ago
- GLM-TTS: Controllable & Emotion-Expressive Zero-shot TTS with Multi-Reward Reinforcement Learning☆1,051Apr 10, 2026Updated 4 months ago
- A TTS that fits in your CPU (and pocket)☆8,705Updated this week
- The open-source AI voice studio. Clone, dictate, create.☆50,755Aug 9, 2026Updated last week
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆22,809May 25, 2026Updated 2 months ago
- ☆217Jun 2, 2026Updated 2 months ago
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching☆1,039Dec 2, 2025Updated 8 months ago
- "ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"☆12,016Jul 29, 2026Updated 3 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling☆210Jun 6, 2026Updated 2 months ago
- MOSS-Speech is a true speech-to-speech large language model without text guidance.☆139Feb 13, 2026Updated 6 months ago
- SoTA open-source TTS☆26,040Jul 21, 2026Updated 3 weeks ago
- A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics…☆967Apr 9, 2026Updated 4 months ago
- An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System☆23,140Updated this week
- MOSS-Transcribe-Diarize 0.9B is an open-source SOTA end-to-end audio understanding model for long-form multi-speaker transcription, diari…☆1,522Jul 24, 2026Updated 3 weeks ago
- super expressive prompting model based on ltx2.3☆475May 23, 2026Updated 2 months ago
- [ACL 2026 Main] Training, inference, and testing of the SAC speech codec model.☆109Nov 1, 2025Updated 9 months ago
- On-device TTS model by Neuphonic☆6,239Jul 30, 2026Updated 2 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation☆452Nov 27, 2025Updated 8 months ago
- Zonos2 is a leading open-weight text-to-speech MoE.☆298Jul 6, 2026Updated last month
- Open source voice AI platform. Self-hosted alternative to Vapi and Retell. On Prem, BYOK across Speech to Speech or LLM/STT/TTS, with a …☆5,402Updated this week
- MiMo-Audio: Audio Language Models are Few-Shot Learners☆1,076Jun 17, 2026Updated 2 months ago
- Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model.☆9,134Updated this week
- 🌋LavaSR: Fast Speech restoration and enhancement☆574Jun 19, 2026Updated 2 months ago
- Code for 'JUST-DUB-IT: Video Dubbing via Joint Audio-Visual Diffusion'☆267May 11, 2026Updated 3 months ago