☆21Jan 28, 2026Updated 5 months ago
Alternatives and similar repositories for J-HARD-TTS-Eval
Users that are interested in J-HARD-TTS-Eval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Quasi-Periodic WaveNet Pytorch implementation☆13Mar 27, 2021Updated 5 years ago
- Coco-Nut (Corpus of connecting NIHONGO utterance and text) corpus☆21Jun 12, 2024Updated 2 years ago
- Unofficial implementation of ConvNeXt-TTS powered by lightning☆18Oct 20, 2024Updated last year
- A repository of Japanese Phoneme-Level BERT☆24Dec 16, 2023Updated 2 years ago
- DDPM-based Pitch Generation and Pitch Controllable Voice Synthesis.☆55Sep 25, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆29Jul 3, 2026Updated 2 weeks ago
- a Frontier Japanese Speech Generation net☆65May 15, 2025Updated last year
- JATTS: A modern, research-oriented Japanese Text-to-speech Open-sourced Toolkit☆43Mar 13, 2026Updated 4 months ago
- Digital Speech Processing in PyTorch.☆15Aug 12, 2022Updated 3 years ago
- JAX implementation of Large Language Models. You can train GPT-2-like model with 青空文庫 (aozora bunko-clean dataset) or any other text dat…☆13Aug 5, 2024Updated last year
- LibriTTS-P: A Corpus with Speaking Style and Speaker Identity Prompts for Text-to-Speech and Style Captioning☆161Jun 13, 2024Updated 2 years ago
- 青空文庫振り仮名注釈付き音声コーパスのデータセット☆50Mar 7, 2025Updated last year
- VAE modified from Descript Audio Codec, which replaces the RVQ with VAE☆92Apr 2, 2024Updated 2 years ago
- LIGHTVOC AN UPSAMPLING-FREE GAN VOCODER BASED ON CONFORMER AND INVERSE SHORT-TIME FOURIER TRANSFORM☆18May 17, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- PromptTTS++: Controlling Speaker Identity in Prompt-Based Text-To-Speech Using Natural Language Descriptions☆86Oct 11, 2024Updated last year
- Provides simple decoding and encoding of audio codecs for Unity.☆17Mar 21, 2023Updated 3 years ago
- Ultra-low-bitrate Speech Codec for Speech Language Modeling Applications☆92Dec 20, 2024Updated last year
- Quasi-Periodic Parallel WaveGAN Pytorch implementation☆46Oct 29, 2022Updated 3 years ago
- Survey of audio language models☆65Apr 18, 2026Updated 3 months ago
- Fine-tuning Moshi/J-Moshi on your own spoken dialogue data☆101Jan 5, 2026Updated 6 months ago
- ☆19Jan 19, 2026Updated 6 months ago
- ☆110Updated this week
- sampling frequency independent convolution for MOS prediction☆15Jul 22, 2025Updated 11 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Layer-wise analysis of self-supervised pre-trained speech representations☆135Oct 18, 2024Updated last year
- Implementation of TTS model based on NVIDIA P-Flow TTS Paper☆77Jul 13, 2026Updated last week
- ☆26Aug 8, 2024Updated last year
- Japanese Converter Kanji to Hiragana, Katakana, Roma-ji☆13Jul 19, 2023Updated 3 years ago
- semantic tokenizer for speech and music☆20Jul 6, 2025Updated last year
- The open source code for SimpleSpeech series☆147Oct 8, 2024Updated last year
- PyTorch implementation of simplified neural source filter model (s-nsf)☆14Aug 4, 2021Updated 4 years ago
- みんなが見たアニメ一覧をまとめて見れるやつ☆11Nov 19, 2025Updated 8 months ago
- [INTERSPEECH 2026] Pre-training, SFT, DPO and GRPO for Text-to-Audio Generation☆48Apr 17, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆65Jan 8, 2025Updated last year
- Alignment files of LibriTTS.☆70Mar 16, 2020Updated 6 years ago
- LightHuBERT: Lightweight and Configurable Speech Representation Learning with Once-for-All Hidden-Unit BERT☆73Sep 26, 2022Updated 3 years ago
- UTokyo-SaruLab MOS Prediction System☆353Apr 2, 2026Updated 3 months ago
- Open-source text-to-speech model from KRAFTON trained exclusively on public speech data, with curated datasets and reproducible training …☆74May 21, 2026Updated 2 months ago
- ☆31Apr 4, 2018Updated 8 years ago
- Speech enhancement in noisy and reverberant environments using deep neural networks☆23Oct 10, 2025Updated 9 months ago