Clone a voice in 5 seconds to generate arbitrary speech in real-time
β60,072Mar 9, 2026Updated 4 months ago
Alternatives and similar repositories for Real-Time-Voice-Cloning
Users that are interested in Real-Time-Voice-Cloning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ45,853Aug 16, 2024Updated last year
- DeepFaceLab is the leading software for creating deepfakes.β19,298Nov 13, 2024Updated last year
- πClone a voice in 5 seconds to generate arbitrary speech in real-timeβ36,919Mar 3, 2026Updated 5 months ago
- Deepfakes Software For Allβ57,235Updated this week
- π Text-Prompted Generative Audio Modelβ39,217Aug 19, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Instant voice cloning by MIT and MyShell. Audio foundation model.β37,086Apr 19, 2025Updated last year
- Real-time face swap for PC streaming or video callsβ31,014Nov 8, 2024Updated last year
- A python package to analyze and compare voices with deep learningβ3,293Oct 12, 2023Updated 2 years ago
- Robust Speech Recognition via Large-Scale Weak Supervisionβ106,614Jul 28, 2026Updated last week
- Stable Diffusion web UIβ164,387Mar 2, 2026Updated 5 months ago
- π€ Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal modelβ¦β163,335Updated this week
- Deep learning for Text to Speech (Discussion forum: https://discourse.mozilla.org/c/tts)β10,168Nov 9, 2023Updated 2 years ago
- The world's simplest facial recognition api for Python and the command lineβ56,657Jun 25, 2026Updated last month
- DeepSpeech is an open source embedded (offline, on-device) speech-to-text engine which can run in real time on devices ranging from a Rasβ¦β26,772Jun 19, 2025Updated last year
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Hunt down social media accounts by username across social networksβ88,212Updated this week
- Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressorβ¦β23,537Mar 3, 2026Updated 5 months ago
- Command-line program to download videos from YouTube.com and other video sitesβ140,867Feb 19, 2026Updated 5 months ago
- Deezer source separation library including pretrained models.β28,352Jun 18, 2026Updated last month
- A latent text-to-image diffusion modelβ73,276Jun 18, 2024Updated 2 years ago
- Avatars for Zoom, Skype and other video-conferencing apps.β16,512Aug 30, 2024Updated last year
- A multi-voice TTS system trained with an emphasis on qualityβ14,866Nov 19, 2024Updated last year
- AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus oβ¦β185,811Updated this week
- This repository contains the codes of "A Lip Sync Expert Is All You Need for Speech to Lip Generation In the Wild", published at ACM Multβ¦β13,146Jun 22, 2025Updated last year
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- This repository contains the source code for the paper First Order Motion Model for Image Animationβ15,009Nov 14, 2024Updated last year
- Tacotron 2 - PyTorch implementation with faster-than-realtime inferenceβ5,298Jun 12, 2024Updated 2 years ago
- SOTA Open Source TTSβ31,987Updated this week
- Making large AI models cheaper, faster and more accessibleβ41,432Jul 13, 2026Updated 3 weeks ago
- End-to-End Speech Processing Toolkitβ9,909Updated this week
- Open source home automation that puts local control and privacy first.β89,686Updated this week
- WaveRNN Vocoder + TTSβ2,191Jul 2, 2022Updated 4 years ago
- openpilot is an operating system for robotics. Currently, it upgrades the driver assistance system on 300+ supported cars.β63,320Updated this week
- An open-source remote desktop application designed for self-hosting, as an alternative to TeamViewer.β119,588Updated this week
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Industry leading face manipulation platformβ29,504Updated this week
- GPT4All: Run Local LLMs on Any Device. Open-source and available for commercial use.β77,412May 27, 2025Updated last year
- All Algorithms implemented in Pythonβ223,473Updated this week
- The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.β123,595Updated this week
- A collective list of free APIsβ454,346Updated this week
- 1 min voice data can also be used to train a good TTS model! (few shot voice cloning)β60,438Jul 22, 2026Updated 2 weeks ago
- Build smaller, faster, and more secure desktop and mobile applications with a web frontend.β109,887Updated this week