Instant voice cloning by MIT and MyShell. Audio foundation model.
β37,066Apr 19, 2025Updated last year
Alternatives and similar repositories for OpenVoice
Users that are interested in OpenVoice are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ45,844Aug 16, 2024Updated last year
- SOTA Open Source TTSβ31,894Updated this week
- π Text-Prompted Generative Audio Modelβ39,218Aug 19, 2024Updated last year
- High-quality multi-lingual text-to-speech library by MyShell.ai. Support English, Spanish, French, Chinese, Japanese and Korean.β7,554Dec 24, 2024Updated last year
- 1 min voice data can also be used to train a good TTS model! (few shot voice cloning)β60,295Jul 22, 2026Updated last week
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Zero-Shot Speech Editing and Text-to-Speech in the Wildβ8,510May 30, 2026Updated 2 months ago
- A generative speech model for daily dialogue.β39,717Apr 10, 2026Updated 3 months ago
- Amphion (/Γ¦mΛfaΙͺΙn/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junioβ¦β9,975Mar 25, 2026Updated 4 months ago
- Inference and training library for high-quality TTS models.β5,582Dec 10, 2024Updated last year
- A coding agent for open models like Kimi K3β67,483Updated this week
- Industry leading face manipulation platformβ29,478Updated this week
- π€― LobeHub is your Chief Agent Operator, organizing your agents into 7Γ24 operations by hiring, scheduling, and reporting on your entire β¦β81,036Updated this week
- Open-Sora: Democratizing Efficient Video Production for Allβ29,228Apr 9, 2026Updated 3 months ago
- Jan is an open source alternative to ChatGPT that runs 100% offline on your computer.β43,800Updated this week
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Robust Speech Recognition via Large-Scale Weak Supervisionβ106,363Updated this week
- Clone a voice in 5 seconds to generate arbitrary speech in real-timeβ60,070Mar 9, 2026Updated 4 months ago
- real time face swap and one-click video deepfake with only a single imageβ95,456Updated this week
- π OpenHands: AI-Driven Developmentβ82,780Updated this week
- The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.β123,112Updated this week
- Foundational model for human-like, expressive TTSβ4,203Jul 30, 2024Updated 2 years ago
- Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.β177,505Updated this week
- Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressorβ¦β23,532Mar 3, 2026Updated 4 months ago
- Drop in a screenshot and convert it to clean code (HTML/Tailwind/React/Vue)β73,789Updated this week
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"β15,053Jul 23, 2026Updated last week
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.β22,527May 25, 2026Updated 2 months ago
- StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Modelsβ6,326Aug 10, 2024Updated last year
- OpenUI let's you describe UI using your imagination, then see it rendered live.β22,492Updated this week
- EmotiVoice π: a Multi-Voice and Prompt-Controlled TTS Engineβ8,507Aug 13, 2024Updated last year
- OCR, layout analysis, reading order, table recognition in 90+ languagesβ21,192Jul 23, 2026Updated last week
- The first real AI developerβ33,718Jun 18, 2026Updated last month
- Build, run, and manage agent platforms.β41,532Updated this week
- Opiniated RAG for integrating GenAI in your apps π§ Focus on your product rather than the RAG. Easy integration in existing products wiβ¦β39,372Jul 9, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- PhotoMaker [CVPR 2024]β10,098Oct 31, 2024Updated last year
- Foundational Models for State-of-the-Art Speech and Text Translationβ11,829Updated this week
- Real-time face swap for PC streaming or video callsβ31,010Nov 8, 2024Updated last year
- Stable Diffusion web UIβ164,326Mar 2, 2026Updated 4 months ago
- A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phoneβ26,071Updated this week
- A sound cloning tool with a web interface, using your voice or any sound to record audio / δΈδΈͺεΈ¦webηι’ηε£°ι³ε ιε·₯ε ·οΌδ½Ώη¨δ½ ηι³θ²ζδ»»ζε£°ι³ζ₯ε½εΆι³ι’β8,989Aug 29, 2025Updated 11 months ago
- Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experienceβ64,194Updated this week