π¬ "Realtime" voice transcription and cloning using ElevenLabs's API.
β56Mar 1, 2023Updated 3 years ago
Alternatives and similar repositories for rtvc
Users that are interested in rtvc are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Speech to text to speech using Elevenlabsβ27Jul 2, 2023Updated 3 years ago
- Listen, transcribe, reply - Voice Assistant using OpenAI & ElevenLabs API'sβ14Jun 24, 2023Updated 3 years ago
- A user-friendly interface for ElevenLabs' API with added audio transcription capability.β13Jun 20, 2023Updated 3 years ago
- This chatbot lets you use your microphone to communicate with GPT-4. It uses the OpenAI text to speech to respond with a voice. It uses Pβ¦β56Dec 6, 2023Updated 2 years ago
- Code for "Distribution-based Emotion Recognition in Conversation"β18Feb 6, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A simple unofficial Python3 library to interface with elevenlabs.io.β17Nov 12, 2023Updated 2 years ago
- Project for HIDING SPEAKERβS SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINEβ15Nov 30, 2022Updated 3 years ago
- A fast CPU-first video/audio transcriber for generating caption files with Whisper and CTranslate2, hosted on Hugging Face Spaces.β11Updated this week
- A diffusion-based cross-lingual voice conversion model, as my bachelor's thesisβ45Jul 24, 2023Updated 3 years ago
- Avocodo: Generative Adversarial Network for Artifact-free Vocoderβ122Jul 14, 2022Updated 4 years ago
- PyTorch implementation of Retriever: Learning Content-Style Representationβ12Jan 27, 2023Updated 3 years ago
- CML-TTS: A Multilingual Dataset for Speech Synthesisβ36Jul 31, 2024Updated 2 years ago
- β11May 7, 2022Updated 4 years ago
- Autonomous and goal-seeking coding agents. π«β22Dec 11, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Google's TPGST reimplementation.β34Dec 11, 2019Updated 6 years ago
- Aty-TTS: Improving fairness for spoken language understanding in atypical speech with Text-to-Speechβ12May 14, 2025Updated last year
- εη¬η»΄ζ€ηδΈζTTSβ34Oct 28, 2022Updated 3 years ago
- Obsidian theme inspired by iA Writerβ16Apr 12, 2024Updated 2 years ago
- NU-Wave 2: A General Neural Audio Upsampling Model for Various Sampling Rates [WIP]β25Jul 5, 2022Updated 4 years ago
- Code for "Phoneme Segmentation Using Self-Supervised Speech Models", Strgar & Harwath, Proceedings of the IEEE Spoken Language Technologyβ¦β55Nov 4, 2022Updated 3 years ago
- β11Aug 7, 2021Updated 5 years ago
- TTSεοΌζζ¬ζ εεοΌε°ζ°εεζ―ε€η转εδΈΊζ±εβ12Apr 27, 2024Updated 2 years ago
- Unofficial Go client for the ElevenLabs API: text-to-speech, speech-to-text, sound generation, and voice management.β68Aug 22, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- β37May 8, 2021Updated 5 years ago
- Syllable Segmentation and Cross-Lingual Generalization in a Visually Grounded, Self-Supervised Speech Modelβ35Aug 27, 2023Updated 3 years ago
- [IJCAI'23] Learning to Speak from Text for Low-Resource TTSβ65May 30, 2023Updated 3 years ago
- ICASSP2022 TTS&VC Summaryβ13Jun 9, 2022Updated 4 years ago
- The source code for the paper CrossSinger (asru2023)β18Oct 12, 2023Updated 2 years ago
- Base mechβ40Updated this week
- python wrap for hts engineβ14Jan 30, 2018Updated 8 years ago
- Rich Prosody Diversity Modelling with Phone-level Mixture Density Networkβ45Dec 1, 2021Updated 4 years ago
- Code for ICASSP 2019 paperβ18Oct 29, 2018Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- https://facetimeanyone.com/β10Nov 9, 2023Updated 2 years ago
- An unofficial PyTorch implementation of Mix-Phoneme-Bertβ40Jul 10, 2023Updated 3 years ago
- MultiSpeaker Tacotron2 using LifeLong Learning.β13Sep 27, 2019Updated 6 years ago
- Finally, some decent sample sentencesβ24Dec 3, 2023Updated 2 years ago
- Manage your Youtube and Twitter subscriptions into groups and foldersβ14May 22, 2026Updated 3 months ago
- Demo audio of VARA-TTS modelβ20Jun 11, 2021Updated 5 years ago
- Audio Generation model working with GPT-2 and VQVAE compressed representation of MelSpectrogramsβ18Oct 8, 2023Updated 2 years ago