🐸Coqui Dialogue Audio Pack contains more than 2000 audio files of synthetic human voices over dialogue created specifically for video games. The pack includes both male and female voices from >30 different voices, and all of the files can be used for commercial purposes (royalty free).
☆47Mar 7, 2023Updated 3 years ago
Alternatives and similar repositories for coqui-voice-pack
Users that are interested in coqui-voice-pack are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Project for HIDING SPEAKER’S SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINE☆15Nov 30, 2022Updated 3 years ago
- Coqui AI TTS plugin☆85Jul 2, 2025Updated last year
- Aty-TTS: Improving fairness for spoken language understanding in atypical speech with Text-to-Speech☆12May 14, 2025Updated last year
- Windows Forms user interface for making lip sync videos with DINet and OpenFace☆26Oct 14, 2023Updated 2 years ago
- 🫠 check your data, before you wreck your model☆16Aug 11, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A pipeline to generate user-preferred photo-realistic avatars using stable-diffusion and bayesian-optimization.☆18Jun 12, 2026Updated 2 months ago
- The YouTube Text-To-Speech dataset is comprised of waveform audio extracted from YouTube videos alongside their English transcriptions☆53Apr 1, 2021Updated 5 years ago
- Syllable Segmentation and Cross-Lingual Generalization in a Visually Grounded, Self-Supervised Speech Model☆35Aug 27, 2023Updated 3 years ago
- Fast trigram-indexed regex search for codebases — 2-6x faster than ripgrep☆21Mar 24, 2026Updated 5 months ago
- ☆12Aug 22, 2017Updated 9 years ago
- Foundational Models for State-of-the-Art Speech and Text Translation☆11Sep 13, 2023Updated 2 years ago
- Training Models Daily☆16Dec 19, 2023Updated 2 years ago
- ☆10Nov 19, 2023Updated 2 years ago
- Basics™ Station Packet Forward protocol using Docker☆15Nov 24, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Upscale images or videos using ESRGAN. TL;DR Converts low res to high res image quality.☆10Aug 4, 2022Updated 4 years ago
- ☆16Oct 4, 2025Updated 11 months ago
- plugin manager for OpenVoiceOS , STT/TTS/Wakewords that can be used anywhere☆14Updated this week
- List of papers about TTS / Список статей о TTS☆10Dec 16, 2017Updated 8 years ago
- Generate audio datasets for training Text-To-Speech models, through smart audio splitting with silence detection, and transcription using…☆30May 27, 2023Updated 3 years ago
- Linguistic processing for Common Voice☆59Jan 18, 2024Updated 2 years ago
- Using Tacotron2 to do Cherokee Text to Speech☆10May 10, 2022Updated 4 years ago
- Deploy a "chat with your data" bot in minutes.☆20Mar 16, 2024Updated 2 years ago
- Open models for Coqui STT☆153May 9, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- TTS Client for Coqui TTS server☆13Jan 7, 2023Updated 3 years ago
- Deep Neural Pitch Extractor for Voice Conversion and TTS Training☆154Aug 22, 2022Updated 4 years ago
- Text to speech is an emerging zone of AI. This repository helps to create a dataset with audio and transcripts for personalized text to s…☆28Mar 14, 2023Updated 3 years ago
- Cherokee Audio data☆11Dec 24, 2023Updated 2 years ago
- A PCB to neatly package a rp2350b with two 12bit high speed adc and four channel audio input for use with hsdaoh☆19May 10, 2025Updated last year
- A python tool that converts Arabic diacritised text to a sequence of phonemes and creates a pronunciation dictionary. This code is based …☆15Sep 5, 2017Updated 8 years ago
- Workflow for forced alignment between languages☆25May 7, 2026Updated 3 months ago
- Reimplementation of Miipher☆30Aug 16, 2023Updated 3 years ago
- Archived origin of secryst/secryst-train/arabic — full history merged there.☆13Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Render wav and convert it with [Diff-SVC](https://github.com/prophesier/diff-svc) model☆10Aug 24, 2025Updated last year
- Googleの音声復元モデルMiipher-2の再現実装の学習および推論コード。学習済みモデルも公開しています。☆33Feb 7, 2026Updated 6 months ago
- Proposed splits for the LREC Wikipron paper☆15Apr 7, 2020Updated 6 years ago
- llmon-py is a multimodal webui for Llama 3-8B.☆16Jul 1, 2024Updated 2 years ago
- Deep learning for thai romanization.☆14Jul 30, 2022Updated 4 years ago
- Arabic Phonetic Dictionary Generator Tool for Automatic Speech Recognition Applications☆11Oct 27, 2021Updated 4 years ago
- ☆11Oct 8, 2023Updated 2 years ago