Mel spectrum based on tacotron2 for melgan speech synthesis
☆15Mar 24, 2023Updated 3 years ago
Alternatives and similar repositories for tacotron2-melgan
Users that are interested in tacotron2-melgan are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A TensorFlow implementation of Google's Tacotron speech synthesis with pre-trained model (unofficial)☆14May 19, 2021Updated 5 years ago
- Real-time melgan based on cpu !!!☆13Dec 3, 2019Updated 6 years ago
- ☆15May 8, 2021Updated 5 years ago
- ICASSP 2021 accepted papers in term of voice conversion (VC)☆18Apr 11, 2021Updated 5 years ago
- ☆31Nov 7, 2018Updated 7 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Using VAEs to do clustering for classification☆11Nov 5, 2017Updated 8 years ago
- ☆19Feb 2, 2023Updated 3 years ago
- Lightweight speaker anonymization [IEEE SLT2021]☆27Jun 6, 2022Updated 4 years ago
- follow NVIDIA, simplify it and support data parallel.☆13Sep 26, 2019Updated 6 years ago
- Converts Mandarin Chinese pinyin notation to IPA (international phonetic alphabet) notation☆19Nov 28, 2023Updated 2 years ago
- A NVIDIA's Pytorch Tacotron2 adaptation with unsupervised Global Style Tokens. The model has been trained with the English read-speech LJ…☆10Sep 4, 2023Updated 2 years ago
- Transfer Learning from Monolingual ASR to Transcription-free Cross-lingual Voice Conversion☆40Oct 22, 2022Updated 3 years ago
- EMPHASIS: An Emotional Phoneme-based Acoustic Model for Speech Synthesis System☆15Mar 31, 2019Updated 7 years ago
- Tacotron2 with Global Style Tokens☆64Apr 19, 2019Updated 7 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Audio Generation model working with GPT-2 and VQVAE compressed representation of MelSpectrograms☆18Oct 8, 2023Updated 2 years ago
- Framework for one-shot multispeaker system based on Deep Learning☆19May 30, 2021Updated 5 years ago
- VAE Tacotron 2, an alternative of GST Tacotron☆91Jul 6, 2023Updated 3 years ago
- The source code for the paper CrossSinger (asru2023)☆18Oct 12, 2023Updated 2 years ago
- ☆37May 8, 2021Updated 5 years ago
- ERISHA is a mulitilingual multispeaker expressive speech synthesis framework. It can transfer the expressivity to the speaker's voice for…☆44Dec 17, 2020Updated 5 years ago
- (R&D) Text to speech using phonemes as inputs and audio codec codes as outputs. Loosely based on MegaByte, VALL-E and Encodec.☆48Sep 4, 2023Updated 2 years ago
- Tensorflow implementation of Chinese/Mandarin TTS (Text-to-Speech) based on Tacotron-2 model.☆132Jul 6, 2023Updated 3 years ago
- Tensorflow Implementation of WaveGlow☆37May 4, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implementation of the Rhythm Formant Analysis methodology for identifying speech rhythms and rhythm variation in the low frequency spectr…☆17Apr 27, 2023Updated 3 years ago
- This paper has been accepted in ACM ICMR 2021.☆20Nov 17, 2025Updated 8 months ago
- Dual-Adversarial Domain Adaptation for replay spoofing detection in automatic speaker verification.☆19Jul 17, 2026Updated last week
- Source code and demo for INTERPSEECH 2023 paper: DuTa-VC: A Duration-aware Typical-to-atypical Voice Conversion Approach with Diffusion P…☆38Dec 5, 2023Updated 2 years ago
- ChiNese Text Normalization (CNTN) tool for Text-to-speech system☆37Apr 12, 2018Updated 8 years ago
- Please visit: https://thuhcsi.github.io/icassp2021-emotion-tts/☆34Mar 17, 2023Updated 3 years ago
- Wavenet pytorch implementation for text-to-speech☆19Jul 19, 2023Updated 3 years ago
- PyTorch implementation of A Neural Algorithm of Artistic Style☆10Dec 20, 2019Updated 6 years ago
- Scaled diffusion transformer for text-to-speech synthesis (DiT + T5Gemma2 conditioning, TorchTitan & Megatron backends, tested up to 1024…☆24Mar 29, 2026Updated 4 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- TTS framework integrating state of the art open source methods (2018/2019)☆48Jun 9, 2026Updated last month
- ☆22Sep 24, 2018Updated 7 years ago
- JAVA 算法数据结构代码 演习实践☆14Jan 5, 2023Updated 3 years ago
- ☆17Aug 27, 2025Updated 11 months ago
- a groory spider .☆12Jul 15, 2017Updated 9 years ago
- 基于scrapy的音频网站爬取☆12Nov 11, 2016Updated 9 years ago
- A ROS package that uses Google Cloud Speech to provide speech to text service☆11Oct 27, 2019Updated 6 years ago