Text-to-Speech Recipe Users can create speech signals from an input text by using text-to-speech (TTS), also referred to as speech synthesis. Popular TTS and Vocoder models, such as Tacotron 2, are supported by SpeechBrain (e.g, HiFIGAN).
☆19Dec 16, 2024Updated last year
Alternatives and similar repositories for Speech
Users that are interested in Speech are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Create speaker voiceprints from a few seconds of audio. And, identify individuals in real-time streaming or recorded conversations.☆17Feb 4, 2019Updated 7 years ago
- Conversion of audio files to text using whisper from OpenAI with a simple tkinter GUI☆10Apr 13, 2023Updated 3 years ago
- ☆10Jul 1, 2019Updated 7 years ago
- A set of Python client examples and utils for https://github.com/carla-simulator/carla☆12Dec 10, 2019Updated 6 years ago
- 無料で使える中品質なテキスト読み上げソフトウェア、VOICEVOXの音声合成エンジン☆10Jan 30, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Topic supervised non-negative matrix factorization with sparse matrices☆12Mar 24, 2020Updated 6 years ago
- ☆10Dec 10, 2021Updated 4 years ago
- JS Effect for REAPER☆14Sep 8, 2025Updated 10 months ago
- Python scripts for AI voice changers☆14Apr 25, 2023Updated 3 years ago
- A auto normalization program for audio followed EBU R128 standard☆10Jan 3, 2021Updated 5 years ago
- ☆41May 15, 2023Updated 3 years ago
- ESPNet TTS with Streamlit GUI☆14Apr 30, 2023Updated 3 years ago
- ☆10Dec 28, 2020Updated 5 years ago
- Notionに毎日新しいarXiv論文のアブストラクト日本語訳 + αを表示するスクリプト☆12Jan 22, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- DeepND ASD & ID☆13Apr 29, 2022Updated 4 years ago
- ☆18Apr 26, 2024Updated 2 years ago
- ☆17Jul 21, 2022Updated 4 years ago
- A simple cross platform graphical user interface (GUI) wrapper to launch executable desktop applications☆17Nov 1, 2024Updated last year
- Code for CVPR19 paper "Monocular Total Capture: Posing Face, Body and Hands in the Wild"☆12May 14, 2021Updated 5 years ago
- Create GUI wrapper for any command line tool☆20Jul 5, 2026Updated 3 weeks ago
- VOICEVOX公式サイトを作るコードです☆14Updated this week
- AI based, large scale, HD material acquisition and creation from a couple pictures.☆11Nov 2, 2021Updated 4 years ago
- This is an official repository for the Article Generation app using Llama2, Pexels, and Streamlit.☆13Aug 5, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- PowerShell GUI application.☆19Jan 12, 2024Updated 2 years ago
- Real-time Speech Separation, Noise Suppression & Speaker Recognition☆17Apr 17, 2019Updated 7 years ago
- transfers weights/colors/positions from one mesh to the other (with arbitrary topology)☆12Jun 12, 2019Updated 7 years ago
- ☆20Jan 7, 2024Updated 2 years ago
- ☆14Oct 11, 2024Updated last year
- This repository is the official implementation of ICASSP2024 paper: Highlight removal network based on an improved dichromatic reflection…☆14Apr 18, 2024Updated 2 years ago
- ☆15Sep 1, 2023Updated 2 years ago
- Author's implementation of learning virtual chimeras by dynamic motion reassembly (SIGGRAPH Asia 2022 Technical Paper)☆16Feb 20, 2023Updated 3 years ago
- ☆10Aug 14, 2020Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 自分の声で音声合成☆17Mar 4, 2019Updated 7 years ago
- Colaboratory notebooks☆13Sep 10, 2020Updated 5 years ago
- Phase Rotation LV2 plugin☆24Apr 19, 2026Updated 3 months ago
- Hosting library for JSFX with CLAP support☆19Oct 10, 2022Updated 3 years ago
- Keras version of Realtime Multi-Person Pose Estimation project☆16Nov 29, 2018Updated 7 years ago
- Open Source bro of my conversational agent. https://github.com/mfmezger/conversational-agent-langchain☆14Feb 5, 2024Updated 2 years ago
- Realistic gramophone noise synthesis using a diffusion model☆18Aug 28, 2022Updated 3 years ago