ttslearn: Library for Pythonで学ぶ音声合成 (Text-to-speech with Python)
☆269Mar 7, 2023Updated 3 years ago
Alternatives and similar repositories for ttslearn
Users that are interested in ttslearn are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- HTS-style full-context labels for JSUT v1.1☆51Apr 16, 2021Updated 5 years ago
- 「Pythonで学ぶ音源分離」のソースコード☆176Jul 5, 2021Updated 5 years ago
- ☆233Nov 13, 2023Updated 2 years ago
- ☆38Sep 20, 2022Updated 3 years ago
- ☆97Apr 12, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- context labels and pronunciation data for JSUT corpus☆77Sep 2, 2021Updated 4 years ago
- Library to build speech synthesis systems designed for easy and fast prototyping.☆399Jun 29, 2024Updated 2 years ago
- Neural network-based singing voice synthesis library for research☆748Oct 9, 2023Updated 2 years ago
- Python wrapper for OpenJTalk☆255Apr 8, 2025Updated last year
- A fork of open_jtalk☆72Mar 31, 2025Updated last year
- UT-Sarulab MOS prediction system using SSL models☆310Apr 11, 2024Updated 2 years ago
- Colaboratory notebooks☆13Sep 10, 2020Updated 5 years ago
- Official implementation of DGP-based multi-speaker speech synthesis with PyTorch☆24Mar 23, 2021Updated 5 years ago
- Python wrapper for Sinsy☆53Oct 9, 2023Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- pyopenjtalk-plus: A Python wrapper for OpenJTalk with additional improvements☆58Updated this week
- WaveGANによる音声生成器☆13Feb 9, 2024Updated 2 years ago
- ☆64May 23, 2022Updated 4 years ago
- The official implementation of VAENAR-TTS, a VAE based non-autoregressive TTS model.☆144Jul 8, 2021Updated 5 years ago
- VITSによるテキスト読み上げ器&ボイスチェンジャー☆93Feb 2, 2023Updated 3 years ago
- ITAコーパスの文章リスト☆240Jul 3, 2026Updated last month
- 44100Hz日本語音源に対応した PITS: Variational Pitch Inference for End-to-end Pitch-controllable TTS without External Pitch Predictor です。☆21May 2, 2023Updated 3 years ago
- Official implementation of the source-filter HiFiGAN vocoder☆277Jul 29, 2023Updated 3 years ago
- Easy-to-Use Speech MOS predictors☆364Oct 24, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch☆1,646Apr 22, 2024Updated 2 years ago
- A repository of Japanese Phoneme-Level BERT☆24Dec 16, 2023Updated 2 years ago
- JVS (Japanese versatile speech) コーパスの自作のラベル☆31Apr 11, 2021Updated 5 years ago
- Ono laboratory audio signal processing exercise for beginners.☆19May 10, 2023Updated 3 years ago
- PyTorch Implementation of Google's Parallel Tacotron 2: A Non-Autoregressive Neural TTS Model with Differentiable Duration Modeling☆191Nov 18, 2021Updated 4 years ago
- A python wrapper for Speech Signal Processing Toolkit (SPTK).☆451Jul 16, 2024Updated 2 years ago
- ☆10Dec 10, 2021Updated 4 years ago
- PITS: Variational Pitch Inference for End-to-end Pitch-controllable TTS without External Pitch Predictor☆280Jul 16, 2023Updated 3 years ago
- 論文執筆チェックリスト☆22Jul 3, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 東北きりたん歌唱データベースの最新ラベルデータ☆148May 1, 2021Updated 5 years ago
- Code for evaluating Japanese pretrained models provided by NTT Ltd.☆246Jun 21, 2023Updated 3 years ago
- [APSIPA'22] Exploring Speaker Age Estimation on Different Self-Supervised Learning Models☆14Oct 19, 2022Updated 3 years ago
- PyTorch Implementation of Google Brain's WaveGrad 2: Iterative Refinement for Text-to-Speech Synthesis