MOSS-TTS-Nano is an open-source multilingual tiny speech generation model from MOSI.AI and the OpenMOSS team. With only 0.1B parameters, it is designed for realtime speech generation, can run directly on CPU without a GPU, and keeps the deployment stack simple enough for local demos, web serving, and lightweight product integration.
☆4,191Jul 26, 2026Updated 3 weeks ago
Alternatives and similar repositories for MOSS-TTS-Nano
Users that are interested in MOSS-TTS-Nano are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fi…☆4,005Jul 26, 2026Updated 3 weeks ago
- VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning☆35,816Updated this week
- ☆56May 7, 2026Updated 3 months ago
- High-Quality Voice Cloning TTS for 600+ Languages☆9,220Updated this week
- MOSS-Audio is an open-source foundation model for unified audio understanding, enabling speech, sound, music, captioning, QA, and reasoni…☆639Jun 2, 2026Updated 2 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- MOSS-Audio-Tokenizer is a Causal Transformer-based audio tokenizer built on the CAT architecture. Trained on 3M hours of diverse audio, i…☆249Jun 16, 2026Updated 2 months ago
- MOSS-VL is the core multimodal model series within the OpenMOSS ecosystem, dedicated to visual understanding.☆435Updated this week
- MOSS-TTSD is a spoken dialogue generation model designed for expressive multi-speaker synthesis. It features long-context modeling, flex…☆1,385Jul 26, 2026Updated 3 weeks ago
- Open-Source Frontier Voice AI☆52,873Jul 24, 2026Updated 3 weeks ago
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆22,809May 25, 2026Updated 2 months ago
- MOSS-Transcribe-Diarize 0.9B is an open-source SOTA end-to-end audio understanding model for long-form multi-speaker transcription, diari…☆1,522Jul 24, 2026Updated 3 weeks ago
- An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System☆23,140Updated this week
- SOTA Open Source TTS☆32,249Aug 3, 2026Updated 2 weeks ago
- Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.☆13,691Jul 24, 2026Updated 3 weeks ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆13,007Mar 17, 2026Updated 5 months ago
- Ming-omni-tts: Simple and Efficient Unified Generation of Speech, Music, and Sound with Precise Control☆265Feb 26, 2026Updated 5 months ago
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching☆1,039Dec 2, 2025Updated 8 months ago
- 🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine☆26,902Jun 14, 2026Updated 2 months ago
- MOSS-Speech is a true speech-to-speech large language model without text guidance.☆139Feb 13, 2026Updated 6 months ago
- The open-source AI voice studio. Clone, dictate, create.☆50,755Aug 9, 2026Updated last week
- Unlimited-length talking video generation that supports image-to-video and video-to-video generation☆7,647May 22, 2026Updated 2 months ago
- On-device TTS model by Neuphonic☆6,239Jul 30, 2026Updated 2 weeks ago
- SoulX-Podcast is an inference codebase by the Soul AI team for generating high-fidelity podcasts from text.☆3,521Dec 11, 2025Updated 8 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- X-Voice☆179Aug 5, 2026Updated 2 weeks ago
- ☆1,202Updated this week
- State-of-the-art TTS model under 25MB 😻☆15,365Jun 11, 2026Updated 2 months ago
- ☆217Jun 2, 2026Updated 2 months ago
- ☆565Apr 3, 2026Updated 4 months ago
- "CLI-Anything: Making ALL Software Agent-Native" -- CLI-Hub: https://clianything.cc/☆47,770Updated this week
- GLM-TTS: Controllable & Emotion-Expressive Zero-shot TTS with Multi-Reward Reinforcement Learning☆1,051Apr 10, 2026Updated 4 months ago
- Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime…☆14,241Updated this week
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆19,911Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The agent that grows with you☆232,420Updated this week
- Mano-P: Open-source GUI-VLA agent for edge devices. #1 on OSWorld (specialized, 58.2%). Runs locally on Apple M4 Mac mini/MacBook — no da…☆2,518Jun 25, 2026Updated last month
- MOVA: Towards Scalable and Synchronized Video–Audio Generation☆1,098Updated this week
- 🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!☆74,928Aug 11, 2026Updated last week
- VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription…☆10,061Updated this week
- Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation☆74Jun 16, 2026Updated 2 months ago
- A high-quality rapid TTS voice cloning model that reaches speeds of 150x realtime.☆5,226Jun 5, 2026Updated 2 months ago