☆21Jan 28, 2026Updated 7 months ago
Alternatives and similar repositories for J-HARD-TTS-Eval
Users that are interested in J-HARD-TTS-Eval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Quasi-Periodic WaveNet Pytorch implementation☆13Mar 27, 2021Updated 5 years ago
- Coco-Nut (Corpus of connecting NIHONGO utterance and text) corpus☆21Jun 12, 2024Updated 2 years ago
- Unofficial implementation of ConvNeXt-TTS powered by lightning☆18Oct 20, 2024Updated last year
- A repository of Japanese Phoneme-Level BERT☆24Dec 16, 2023Updated 2 years ago
- DDPM-based Pitch Generation and Pitch Controllable Voice Synthesis.☆55Sep 25, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆30Jul 3, 2026Updated last month
- a Frontier Japanese Speech Generation net☆65May 15, 2025Updated last year
- JATTS: A modern, research-oriented Japanese Text-to-speech Open-sourced Toolkit☆44Mar 13, 2026Updated 5 months ago
- Digital Speech Processing in PyTorch.☆15Aug 12, 2022Updated 4 years ago
- JAX implementation of Large Language Models. You can train GPT-2-like model with 青空文庫 (aozora bunko-clean dataset) or any other text dat…☆13Aug 5, 2024Updated 2 years ago
- LibriTTS-P: A Corpus with Speaking Style and Speaker Identity Prompts for Text-to-Speech and Style Captioning☆165Jun 13, 2024Updated 2 years ago
- VAE modified from Descript Audio Codec, which replaces the RVQ with VAE☆92Apr 2, 2024Updated 2 years ago
- LIGHTVOC AN UPSAMPLING-FREE GAN VOCODER BASED ON CONFORMER AND INVERSE SHORT-TIME FOURIER TRANSFORM☆18May 17, 2024Updated 2 years ago
- PromptTTS++: Controlling Speaker Identity in Prompt-Based Text-To-Speech Using Natural Language Descriptions☆86Oct 11, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Provides simple decoding and encoding of audio codecs for Unity.☆17Mar 21, 2023Updated 3 years ago
- Ultra-low-bitrate Speech Codec for Speech Language Modeling Applications☆92Dec 20, 2024Updated last year
- Quasi-Periodic Parallel WaveGAN Pytorch implementation☆46Oct 29, 2022Updated 3 years ago
- Survey of audio language models☆65Apr 18, 2026Updated 4 months ago
- Fine-tuning Moshi/J-Moshi on your own spoken dialogue data☆105Jan 5, 2026Updated 7 months ago
- ☆19Jan 19, 2026Updated 7 months ago
- sampling frequency independent convolution for MOS prediction☆16Jul 22, 2025Updated last year
- ☆112Jul 15, 2026Updated last month
- Layer-wise analysis of self-supervised pre-trained speech representations☆135Oct 18, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of TTS model based on NVIDIA P-Flow TTS Paper☆77Jul 13, 2026Updated last month
- ☆26Aug 8, 2024Updated 2 years ago
- Japanese Converter Kanji to Hiragana, Katakana, Roma-ji☆13Jul 19, 2023Updated 3 years ago
- semantic tokenizer for speech and music☆20Jul 6, 2025Updated last year
- The open source code for SimpleSpeech series☆146Oct 8, 2024Updated last year
- PyTorch implementation of simplified neural source filter model (s-nsf)☆14Aug 4, 2021Updated 5 years ago
- 青空文庫振り仮名注釈付き音声コーパスのデータセット☆50Mar 7, 2025Updated last year
- みんなが見たアニメ一覧をまとめて見れるやつ☆11Nov 19, 2025Updated 9 months ago
- [INTERSPEECH 2026] Pre-training, SFT, DPO and GRPO for Text-to-Audio Generation☆50Apr 17, 2026Updated 4 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆65Jan 8, 2025Updated last year
- Alignment files of LibriTTS.☆70Mar 16, 2020Updated 6 years ago
- LightHuBERT: Lightweight and Configurable Speech Representation Learning with Once-for-All Hidden-Unit BERT☆74Sep 26, 2022Updated 3 years ago
- UTokyo-SaruLab MOS Prediction System☆365Apr 2, 2026Updated 4 months ago
- Open-source text-to-speech model from KRAFTON trained exclusively on public speech data, with curated datasets and reproducible training …☆97May 21, 2026Updated 3 months ago
- ☆31Apr 4, 2018Updated 8 years ago
- Speech enhancement in noisy and reverberant environments using deep neural networks☆23Oct 10, 2025Updated 10 months ago