A 100M-parameter multilingual TTS model for real-time CPU inference, voice cloning, and 48 kHz stereo generation
☆4,414Sep 6, 2026Updated 3 weeks ago
Alternatives and similar repositories for MOSS-TTS-Nano
Users that are interested in MOSS-TTS-Nano are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An open-source model family for long-form speech, dialogue synthesis, voice design, sound effects, and real-time streaming TTS☆4,142Sep 6, 2026Updated 3 weeks ago
- VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning☆38,017Sep 2, 2026Updated 3 weeks ago
- A local browser reading app powered by MOSS-TTS-Nano with in-browser ONNX inference☆61May 7, 2026Updated 4 months ago
- High-Quality Voice Cloning TTS for 600+ Languages☆13,953Updated this week
- An open-source model for understanding speech, environmental sounds, and music through captioning, question answering, and reasoning☆673Sep 6, 2026Updated 3 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A 1.6B causal Transformer audio tokenizer with streaming, variable bitrates, and semantic alignment across speech, sound, and music☆258Jun 16, 2026Updated 3 months ago
- An open-weight 11B model series for long-form and real-time video understanding☆739Updated this week
- A multilingual model for long-form, multi-speaker dialogue synthesis with flexible speaker control and zero-shot voice cloning☆1,400Sep 6, 2026Updated 3 weeks ago
- A 0.9B model for long-form transcription in 50+ languages with speaker diarization, timestamps, and acoustic event awareness☆2,086Sep 15, 2026Updated last week
- Open-Source Frontier Voice AI☆54,512Sep 3, 2026Updated 3 weeks ago
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆23,775May 25, 2026Updated 4 months ago
- An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System☆24,202Aug 18, 2026Updated last month
- SOTA Open Source TTS☆32,863Sep 16, 2026Updated last week
- Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.☆13,791Sep 9, 2026Updated 2 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆13,562Mar 17, 2026Updated 6 months ago
- Ming-omni-tts: Simple and Efficient Unified Generation of Speech, Music, and Sound with Precise Control☆265Feb 26, 2026Updated 7 months ago
- 🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine☆28,452Jun 14, 2026Updated 3 months ago
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching☆1,074Dec 2, 2025Updated 9 months ago
- An end-to-end speech-to-speech language model that generates spoken responses without text guidance☆139Feb 13, 2026Updated 7 months ago
- The open-source AI voice studio. Clone, dictate, create.☆55,829Aug 9, 2026Updated last month
- Unlimited-length talking video generation that supports image-to-video and video-to-video generation☆7,927May 22, 2026Updated 4 months ago
- On-device TTS model by Neuphonic☆6,289Jul 30, 2026Updated last month
- SoulX-Podcast is an inference codebase by the Soul AI team for generating high-fidelity podcasts from text.☆3,561Dec 11, 2025Updated 9 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- State-of-the-art TTS model under 25MB 😻☆15,488Aug 19, 2026Updated last month
- X-Voice☆183Updated this week
- ☆1,343Aug 17, 2026Updated last month
- ☆218Jun 2, 2026Updated 3 months ago
- "CLI-Anything: Making ALL Software Agent-Native" -- CLI-Hub: https://clianything.cc/☆50,749Updated this week
- ☆579Apr 3, 2026Updated 5 months ago
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆20,517Updated this week
- Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime…☆14,989Updated this week
- GLM-TTS: Controllable & Emotion-Expressive Zero-shot TTS with Multi-Reward Reinforcement Learning☆1,067Apr 10, 2026Updated 5 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- The agent that grows with you☆249,407Updated this week
- Mano-P: Open-source GUI-VLA agent for edge devices. #1 on OSWorld (specialized, 58.2%). Runs locally on Apple M4 Mac mini/MacBook — no da…☆2,793Jun 25, 2026Updated 3 months ago
- A foundation model that generates synchronized video and audio in a single model☆1,119Updated this week
- 🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl! Don't be shy, join here: http…☆84,019Updated this week
- Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation☆75Jun 16, 2026Updated 3 months ago
- A high-quality rapid TTS voice cloning model that reaches speeds of 150x realtime.☆5,411Jun 5, 2026Updated 3 months ago
- A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy injectors to reality checkers. Ea…☆154,869Updated this week