An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
☆22,226Jul 14, 2026Updated 2 weeks ago
Alternatives and similar repositories for index-tts
Users that are interested in index-tts are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆22,464May 25, 2026Updated 2 months ago
- SOTA Open Source TTS☆31,398Updated this week
- Added vLLM support to IndexTTS for faster inference.☆1,214Apr 13, 2026Updated 3 months ago
- 1 min voice data can also be used to train a good TTS model! (few shot voice cloning)☆60,170Jul 22, 2026Updated last week
- VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning☆34,391Jul 8, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A generative speech model for daily dialogue.☆39,689Apr 10, 2026Updated 3 months ago
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆19,514Updated this week
- Spark-TTS Inference Code☆11,001Apr 9, 2025Updated last year
- Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"☆15,039Updated this week
- Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio…☆8,949Updated this week
- The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.☆122,600Updated this week
- Unlimited-length talking video generation that supports image-to-video and video-to-video generation☆7,536May 22, 2026Updated 2 months ago
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆12,648Mar 17, 2026Updated 4 months ago
- ☆6,079Jun 15, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- SoulX-Podcast is an inference codebase by the Soul AI team for generating high-fidelity podcasts from text.☆3,505Dec 11, 2025Updated 7 months ago
- 🚀 Truly open-source AI avatar(digital human) toolkit for offline video generation and digital human cloning.☆14,217Apr 21, 2026Updated 3 months ago
- Open-Source Frontier Voice AI☆50,834Updated this week
- Taming Stable Diffusion for Lip Sync!☆5,931Jun 20, 2025Updated last year
- Text-audio foundation model from Boson AI☆8,302Jun 5, 2026Updated last month
- Translate the video from one language to another and embed dubbing & subtitles.☆18,471Updated this week
- zero-shot voice conversion & singing voice conversion, with real-time support☆3,891Apr 20, 2025Updated last year
- Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self…☆150,563Updated this week
- AI productivity studio with smart chat, autonomous agents, and 300+ assistants. Unified access to frontier LLMs☆49,062Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Long-form streaming TTS system for multi-speaker dialogue generation☆1,416Oct 26, 2025Updated 9 months ago
- 小红书笔记 | 评论爬虫、抖音视频 | 评论爬虫、快手视频 | 评论爬虫、B 站视频 | 评论爬虫、微博帖子 | 评论爬虫、百度贴吧帖子 | 百度贴吧评论回复爬虫 | 知乎问答文章|评论爬虫☆58,431Updated this week
- 利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow.☆99,671Updated this week
- 🎬 卡卡字幕助手 | VideoCaptioner - 基于 LLM 的智能字幕助手 - 视频字幕生成、断句、校 正、字幕翻译全流程处理!- A powered tool for easy and efficient video subtitling.☆15,448Jul 19, 2026Updated last week
- Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.☆75,993Updated this week
- MOSS-TTSD is a spoken dialogue generation model designed for expressive multi-speaker synthesis. It features long-context modeling, flex…☆1,363Updated this week
- GLM-TTS: Controllable & Emotion-Expressive Zero-shot TTS with Multi-Reward Reinforcement Learning☆1,046Apr 10, 2026Updated 3 months ago
- 基于AI的图片/视频硬字幕去除、文本水印去除,无损分辨率生成去字幕、去水印后的图片/视频文件。无需申请第三方API,本地实现。AI-based tool for removing hard-coded subtitles and text-like watermarks f…☆12,095Jun 30, 2026Updated 3 weeks ago
- A unified AI model hub for aggregation & distribution. It supports cross-converting various LLMs into OpenAI-compatible, Claude-compatibl…☆43,668Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 微舆:人人可用的多Agent舆情分析助手,打破信息茧房,还原舆情原貌,预测未来走向,辅助决策!从0实现,不依赖任何框架。☆41,878Jul 21, 2026Updated last week
- A cross-platform desktop All-in-One assistant for Claude Code, Codex, OpenCode, OpenClaw, Grok Build & Hermes Agent. Only official websit…☆121,972Updated this week
- ⭐AI-driven public opinion & trend monitor with multi-platform aggregation, RSS, and smart alerts.🎯 告别信息过载,你的 AI 舆情监控助手与热点筛选工具!聚合多平台热点 + …☆60,969Jul 17, 2026Updated last week
- ☆34Mar 16, 2026Updated 4 months ago
- 使用IndexTTS模型在ComfyUI中实现高质量文本到语音转换的自定义节点。支持中文和英文文本,可以基于参考音频复刻声音特征。☆726Jun 29, 2026Updated last month
- Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime…☆13,837Updated this week
- Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/…☆86,350Jul 22, 2026Updated last week