List of open-source TTS, voice cloning, and music generation models
☆487Sep 10, 2026Updated this week
Alternatives and similar repositories for awesome-ai-voice
Users that are interested in awesome-ai-voice are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Programmatic video for the web. Define videos with a fluent TypeScript API, compile them to a portable JSON format, and render to MP4 — i…☆145Aug 12, 2026Updated last month
- ☆89Jun 20, 2026Updated 2 months ago
- Zero-shot expressive voice cloning and speech generation. Generate anything from short clips to full-length audiobooks with realistic emo…☆551Jul 7, 2026Updated 2 months ago
- Draft to Take beta: local-first AI audio production studio powered by IndexTTS2, Docker, Qwen, OmniVoice, SFX, ambience, and music sideca…☆77Aug 12, 2026Updated last month
- ☆313May 1, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- rotating proxy system☆25Updated this week
- Self-hosted, OpenAI-compatible RAG API + MCP server that plugs local knowledge into existing LLM clients.☆34Updated this week
- Multimodal AI studio powered by Qwen3.6-35B-A3B. End-to-end web app exposing visual reasoning, image captioning, and document understandi…☆28Apr 23, 2026Updated 4 months ago
- Open Source Alternative to Lovable, v0, Bolt, Replit, Emergent. 🌟 Star if you like it!☆281Updated this week
- Watch any social video → get an architecture diagram, working component, runnable notebook, or step-by-step cheat sheet — automatically.☆250Sep 5, 2026Updated last week
- A lightweight, fully-featured TypeScript SDK for the Shopee Open API v2. Features automatic request signing, pluggable token storage, rob…☆198Updated this week
- A local AI image generator and editor powered by open image models.☆114Jun 28, 2026Updated 2 months ago
- ☆63Apr 8, 2026Updated 5 months ago
- Advanced drum machine for ComfyUI featuring a 64-step sequencer, custom sample support, and retro hardware aesthetics.☆20Jul 2, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆11Jul 28, 2025Updated last year
- All available LTX-2 models, encoders, workflows, LoRAs for ComfyUI☆592Updated this week
- X-Voice☆182Aug 5, 2026Updated last month
- 録音不要でオリジナルAI音声の教師データを作るGUIツール☆181Jun 10, 2026Updated 3 months ago
- ☆74May 8, 2026Updated 4 months ago
- 🎵 The Ultimate Open Source Suno Alternative - Professional UI for ACE-Step 1.5 AI Music Generation. Free, local, unlimited. Stop paying …☆4,867Jun 27, 2026Updated 2 months ago
- Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with…☆12,794Jul 13, 2026Updated 2 months ago
- Open-source text-to-speech model from KRAFTON trained exclusively on public speech data, with curated datasets and reproducible training …☆101May 21, 2026Updated 3 months ago
- Auto generate AI video with hyperframes☆370Sep 5, 2026Updated last week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- High-Quality Voice Cloning TTS for 600+ Languages☆12,689Updated this week
- Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality • Chuyển văn bản thành giọng nói ti…☆2,548Updated this week
- Building actual open source including dataset Multilingual TTS more than 150 languages with Voice Cloning.☆57Aug 26, 2026Updated 2 weeks ago
- Native OpenFX retro / analog / glitch / CRT-VHS plugins for DaVinci Resolve☆24Jul 10, 2026Updated 2 months ago
- Local social media automation dashboard for TikTok, Instagram, and YouTube.☆1,116Jun 10, 2026Updated 3 months ago
- Real-time speech translation — macOS & Windows, free TTS, no server, your API keys only☆1,283Jul 11, 2026Updated 2 months ago
- Freelance & agency workspace — CRM, Kanban, contracts, infra renewals, AI search. React 19 + Vite 6 + Tailwind 4. EN/ZH/JA/ES/PT/KO/VI.☆27May 23, 2026Updated 3 months ago
- 138 bilingual AI marketing skills (69 VN + 69 Global) for Claude Code, OpenCode, Codex, VS Code. Four role SOP packs — content, design, p…☆573Updated this week
- The open-source CapCut alternative for Linux and Windows☆302Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆21Jul 15, 2026Updated last month
- [CVPR 2026] PersonaLive! : Expressive Portrait Image Animation for Live Streaming☆3,718Aug 28, 2026Updated 2 weeks ago
- CLI tool that auto-processes video recordings: transcribes, removes silence, generates captions, creates shorts, social posts, and more☆221Updated this week
- Stable Audio LoRA Trainer of salty goodness☆99Sep 3, 2026Updated last week
- Real-time text-to-speech with Qwen3-TTS☆1,347Aug 25, 2026Updated 2 weeks ago
- Zonos2 is a leading open-weight text-to-speech MoE.☆311Jul 6, 2026Updated 2 months ago
- [CVPR 2026] Cross-Modal Emotion Transfer for Emotion Editing in Talking Face Video - Official PyTorch Implementation☆37Aug 20, 2026Updated 3 weeks ago