A 100M-parameter multilingual TTS model for real-time CPU inference, voice cloning, and 48 kHz stereo generation
☆4,291Sep 6, 2026Updated this week
Alternatives and similar repositories for MOSS-TTS-Nano
Users that are interested in MOSS-TTS-Nano are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An open-source model family for long-form speech, dialogue synthesis, voice design, sound effects, and real-time streaming TTS☆4,070Updated this week
- VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning☆36,815Updated this week
- A local browser reading app powered by MOSS-TTS-Nano with in-browser ONNX inference☆60May 7, 2026Updated 4 months ago
- High-Quality Voice Cloning TTS for 600+ Languages☆10,351Aug 31, 2026Updated last week
- An open-source model for understanding speech, environmental sounds, and music through captioning, question answering, and reasoning☆656Updated this week
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- A 1.6B causal Transformer audio tokenizer with streaming, variable bitrates, and semantic alignment across speech, sound, and music☆255Jun 16, 2026Updated 2 months ago
- An open-weight 11B model series for long-form and real-time video understanding☆599Updated this week
- A multilingual model for long-form, multi-speaker dialogue synthesis with flexible speaker control and zero-shot voice cloning☆1,394Updated this week
- Open-Source Frontier Voice AI☆53,885Updated this week
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆23,494May 25, 2026Updated 3 months ago
- A 0.9B model for long-form transcription in 50+ languages with speaker diarization, timestamps, and acoustic event awareness☆1,865Updated this week
- An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System☆23,806Aug 18, 2026Updated 3 weeks ago
- SOTA Open Source TTS☆32,589Aug 22, 2026Updated 2 weeks ago
- Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.☆13,763Jul 24, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆13,303Mar 17, 2026Updated 5 months ago
- Ming-omni-tts: Simple and Efficient Unified Generation of Speech, Music, and Sound with Precise Control☆265Feb 26, 2026Updated 6 months ago
- 🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine☆27,859Jun 14, 2026Updated 2 months ago
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching☆1,054Dec 2, 2025Updated 9 months ago
- An end-to-end speech-to-speech language model that generates spoken responses without text guidance☆140Feb 13, 2026Updated 6 months ago
- The open-source AI voice studio. Clone, dictate, create.☆52,492Aug 9, 2026Updated 3 weeks ago
- Unlimited-length talking video generation that supports image-to-video and video-to-video generation☆7,801May 22, 2026Updated 3 months ago
- On-device TTS model by Neuphonic☆6,272Jul 30, 2026Updated last month
- SoulX-Podcast is an inference codebase by the Soul AI team for generating high-fidelity podcasts from text.☆3,547Dec 11, 2025Updated 8 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- X-Voice☆181Aug 5, 2026Updated last month
- ☆1,316Aug 17, 2026Updated 3 weeks ago
- State-of-the-art TTS model under 25MB 😻☆15,431Aug 19, 2026Updated 2 weeks ago
- ☆216Jun 2, 2026Updated 3 months ago
- ☆572Apr 3, 2026Updated 5 months ago
- "CLI-Anything: Making ALL Software Agent-Native" -- CLI-Hub: https://clianything.cc/☆49,096Aug 21, 2026Updated 2 weeks ago
- GLM-TTS: Controllable & Emotion-Expressive Zero-shot TTS with Multi-Reward Reinforcement Learning☆1,063Apr 10, 2026Updated 4 months ago
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆20,218Updated this week
- Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime…☆14,653Updated this week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The agent that grows with you☆242,948Updated this week
- Mano-P: Open-source GUI-VLA agent for edge devices. #1 on OSWorld (specialized, 58.2%). Runs locally on Apple M4 Mac mini/MacBook — no da…☆2,688Jun 25, 2026Updated 2 months ago
- A foundation model that generates synchronized video and audio in a single model☆1,110Updated this week
- 🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!☆79,037Updated this week
- Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation☆75Jun 16, 2026Updated 2 months ago
- A high-quality rapid TTS voice cloning model that reaches speeds of 150x realtime.☆5,341Jun 5, 2026Updated 3 months ago
- VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription…☆20,621Updated this week