Instant voice cloning by MIT and MyShell. Audio foundation model.
β37,705Apr 19, 2025Updated last year
Alternatives and similar repositories for OpenVoice
Users that are interested in OpenVoice are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ46,075Aug 16, 2024Updated 2 years ago
- SOTA Open Source TTSβ32,863Sep 16, 2026Updated last week
- π Text-Prompted Generative Audio Modelβ39,272Aug 19, 2024Updated 2 years ago
- High-quality multi-lingual text-to-speech library by MyShell.ai. Support English, Spanish, French, Chinese, Japanese and Korean.β7,650Dec 24, 2024Updated last year
- 1 min voice data can also be used to train a good TTS model! (few shot voice cloning)β62,228Aug 18, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Zero-Shot Speech Editing and Text-to-Speech in the Wildβ8,574May 30, 2026Updated 3 months ago
- A generative speech model for daily dialogue.β39,874Apr 10, 2026Updated 5 months ago
- Amphion (/Γ¦mΛfaΙͺΙn/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junioβ¦β10,305Mar 25, 2026Updated 6 months ago
- Inference and training library for high-quality TTS models.β5,591Dec 10, 2024Updated last year
- A coding agent for open models like Kimi K3 and GLM 5.3β68,463Updated this week
- Industry leading face manipulation platformβ30,066Updated this week
- π€― LobeHub is your Chief Agent Operator, organizing your agents into 7Γ24 operations by hiring, scheduling, and reporting on your entire β¦β82,873Updated this week
- Open-Sora: Democratizing Efficient Video Production for Allβ29,850Apr 9, 2026Updated 5 months ago
- Jan is an open source alternative to ChatGPT that runs 100% offline on your computer.β44,673Updated this week
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Robust Speech Recognition via Large-Scale Weak Supervisionβ109,693Aug 31, 2026Updated 3 weeks ago
- Clone a voice in 5 seconds to generate arbitrary speech in real-timeβ60,153Mar 9, 2026Updated 6 months ago
- real time face swap and one-click video deepfake with only a single imageβ96,862Updated this week
- π OpenHands: AI-Driven Developmentβ89,385Updated this week
- The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface. The fastest local inference engine in thβ¦β135,348Updated this week
- Foundational model for human-like, expressive TTSβ4,205Jul 30, 2024Updated 2 years ago
- Get up and running with Kimi, GLM, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.β181,856Updated this week
- Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressorβ¦β23,651Mar 3, 2026Updated 6 months ago
- Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"β15,308Sep 21, 2026Updated last week
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.β23,775May 25, 2026Updated 4 months ago
- Drop in a screenshot and convert it to clean code (HTML/Tailwind/React/Vue)β79,819Sep 9, 2026Updated 2 weeks ago
- StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Modelsβ6,360Aug 10, 2024Updated 2 years ago
- OpenUI let's you describe UI using your imagination, then see it rendered live.β22,572Updated this week
- EmotiVoice π: a Multi-Voice and Prompt-Controlled TTS Engineβ8,531Sep 3, 2026Updated 3 weeks ago
- OCR, layout analysis, reading order, table recognition in 90+ languagesβ21,424Sep 11, 2026Updated 2 weeks ago
- The first real AI developerβ33,660Jun 18, 2026Updated 3 months ago
- Build, run, and manage agent platforms.β42,367Updated this week
- Opiniated RAG for integrating GenAI in your apps π§ Focus on your product rather than the RAG. Easy integration in existing products wiβ¦β39,559Aug 31, 2026Updated 3 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PhotoMaker [CVPR 2024]β10,086Oct 31, 2024Updated last year
- Foundational Models for State-of-the-Art Speech and Text Translationβ11,879Sep 8, 2026Updated 3 weeks ago
- Real-time face swap for PC streaming or video callsβ31,009Nov 8, 2024Updated last year
- Stable Diffusion web UIβ165,143Mar 2, 2026Updated 6 months ago
- A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phoneβ26,469Sep 8, 2026Updated 3 weeks ago
- A sound cloning tool with a web interface, using your voice or any sound to record audio / δΈδΈͺεΈ¦webηι’ηε£°ι³ε ιε·₯ε ·οΌδ½Ώη¨δ½ ηι³θ²ζδ»»ζε£°ι³ζ₯ε½εΆι³ι’β8,968Aug 29, 2025Updated last year
- Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experienceβ66,551Updated this week