Text-to-Speech Recipe Users can create speech signals from an input text by using text-to-speech (TTS), also referred to as speech synthesis. Popular TTS and Vocoder models, such as Tacotron 2, are supported by SpeechBrain (e.g, HiFIGAN).
☆19Dec 16, 2024Updated last year
Alternatives and similar repositories for Speech
Users that are interested in Speech are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Create speaker voiceprints from a few seconds of audio. And, identify individuals in real-time streaming or recorded conversations.☆17Feb 4, 2019Updated 7 years ago
- Conversion of audio files to text using whisper from OpenAI with a simple tkinter GUI☆10Apr 13, 2023Updated 3 years ago
- Speaker Identification using Neural Net.☆20Jul 30, 2024Updated 2 years ago
- MongoDB with Pymongo Tutorial☆10Apr 19, 2024Updated 2 years ago
- A set of Python client examples and utils for https://github.com/carla-simulator/carla☆12Dec 10, 2019Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 無料で使える中品質なテキスト読み上げソフトウェア、VOICEVOXの音声合成エンジン☆10Jan 30, 2023Updated 3 years ago
- Topic supervised non-negative matrix factorization with sparse matrices☆12Mar 24, 2020Updated 6 years ago
- ☆10Dec 10, 2021Updated 4 years ago
- Image Captioning Agent using Mistral 7B☆11Dec 1, 2023Updated 2 years ago
- ☆10Jul 20, 2023Updated 3 years ago
- Python scripts for AI voice changers☆14Apr 25, 2023Updated 3 years ago
- A auto normalization program for audio followed EBU R128 standard☆10Jan 3, 2021Updated 5 years ago
- Using PDFPlumber for PDF data extraction☆15May 31, 2017Updated 9 years ago
- MLOps Engineer Roadmap 2026. This repository collects the best 100% FREE video tutorials and official documentation to help you become a …☆21Jan 2, 2026Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆41May 15, 2023Updated 3 years ago
- Python profiling tool☆17Jan 28, 2025Updated last year
- ☆10Dec 28, 2020Updated 5 years ago
- Notionに毎日新しいarXiv論文のアブストラクト日本語訳 + αを表示するスクリプト☆12Jan 22, 2023Updated 3 years ago
- ☆18Apr 26, 2024Updated 2 years ago
- ☆17Jul 21, 2022Updated 4 years ago
- A simple cross platform graphical user interface (GUI) wrapper to launch executable desktop applications☆17Nov 1, 2024Updated last year
- chinese_tacotron-2☆12Feb 27, 2018Updated 8 years ago
- VOICEVOX公式サイトを作るコードです☆14Aug 19, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This is an introduction to Retrieval-Augmented Generation (RAG) for beginners . It uses Llama 2 LLM, FAISS vector store, and LangChain as…☆17Jul 8, 2025Updated last year
- AI based, large scale, HD material acquisition and creation from a couple pictures.☆11Nov 2, 2021Updated 4 years ago
- This is an official repository for the Article Generation app using Llama2, Pexels, and Streamlit.☆13Aug 5, 2023Updated 3 years ago
- transfers weights/colors/positions from one mesh to the other (with arbitrary topology)☆12Jun 12, 2019Updated 7 years ago
- ☆14Oct 11, 2024Updated last year
- ☆15Sep 1, 2023Updated 2 years ago
- Aqueduct Project blender addon☆11Feb 25, 2021Updated 5 years ago
- Official implementation of Kimodo, a kinematic motion diffusion model for high-quality human(oid) motion generation.☆48Jun 18, 2026Updated 2 months ago
- Author's implementation of learning virtual chimeras by dynamic motion reassembly (SIGGRAPH Asia 2022 Technical Paper)☆16Feb 20, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 自分の声で音声合成☆17Mar 4, 2019Updated 7 years ago
- Phase Rotation LV2 plugin☆24Updated this week
- Keras version of Realtime Multi-Person Pose Estimation project☆16Nov 29, 2018Updated 7 years ago
- Realistic gramophone noise synthesis using a diffusion model☆18Aug 28, 2022Updated 4 years ago
- TF Mesh Renderer☆15Dec 25, 2019Updated 6 years ago
- Implementation of vocoders empowered with pytorch lightning☆18Jan 27, 2024Updated 2 years ago
- [ACL-IJCNLP 2021] "EarlyBERT: Efficient BERT Training via Early-bird Lottery Tickets" by Xiaohan Chen, Yu Cheng, Shuohang Wang, Zhe Gan, …☆18Dec 30, 2021Updated 4 years ago