Onnx compatible styletts2 code
☆16Apr 4, 2026Updated 5 months ago
Alternatives and similar repositories for styletts2-inference
Users that are interested in styletts2-inference are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- StyleTTS2 + Vocos as a Decoder☆13Jul 31, 2026Updated last month
- Tool to make high quality text to speech (tts) corpus from audio + text books.☆27Jul 31, 2025Updated last year
- High quality text-to-speech based on StyleTTS 2.☆79Apr 6, 2026Updated 5 months ago
- StyleTTS 2 Optimized Training Fork☆32Feb 2, 2025Updated last year
- A lightweight, efficient variation of the StyleTTS 2 text‐to‐speech model.☆52May 22, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Export an ONNX graph that performs ISTFT. Designed for TTS models.☆28Apr 23, 2024Updated 2 years ago
- ☆22Apr 29, 2025Updated last year
- ☆17Jun 2, 2025Updated last year
- Simple inference for Vits2 TTS Using ONNXRUNTIME and espeak-ng on C++☆19Apr 17, 2024Updated 2 years ago
- A unified evaluation suite for speech-to-text translation, covering SpeechLLMs, SFMs, and cascaded systems across diverse real-world spee…☆33Aug 9, 2026Updated 3 weeks ago
- ☆12Jun 14, 2024Updated 2 years ago
- ☆53Aug 28, 2024Updated 2 years ago
- built a 124M param GPT☆23Jan 28, 2025Updated last year
- This repository implement a novel zero-shot TTS framework, named Flamed-TTS, focusing on the efficient generation and dynamic pacing in …☆57Aug 9, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A collection of all our phonemeizers for dataset construction and inference☆32Feb 21, 2025Updated last year
- Generate audio datasets for training Text-To-Speech models, through smart audio splitting with silence detection, and transcription using…☆30May 27, 2023Updated 3 years ago
- Official repo for DisCoder: High-Fidelity Music Vocoder using Neural Audio Codecs presented at ICASSP 2025☆45Feb 24, 2025Updated last year
- ☆12Jan 14, 2020Updated 6 years ago
- Code for ACL 2024 main conference paper "Can We Achieve High-quality Direct Speech-to-Speech Translation Without Parallel Speech Data?".☆27Jul 2, 2024Updated 2 years ago
- ☆28Jun 22, 2026Updated 2 months ago
- Simple diarization model☆53Jun 13, 2025Updated last year
- MAGIC-TTS: Fine-Grained Controllable Speech Synthesis with Explicit Local Duration and Pause Control☆55Apr 28, 2026Updated 4 months ago
- openFrameworks implementation of whisper.cpp☆13Sep 14, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Converts CLIP models to ONNX☆11Jan 17, 2023Updated 3 years ago
- Arabic Grapheme-to-Phoneme (G2P) Conversion☆15Mar 15, 2025Updated last year
- Prosody and Pronunciation Modification Network☆64May 5, 2025Updated last year
- used to evaluate wavenet vocoder by rmse f0, MCD, rmse ap...☆15Jan 20, 2020Updated 6 years ago
- LDC: Lightweight Dense CNN for Edge DetectionのPythonでのONNX推論サンプル☆15May 6, 2023Updated 3 years ago
- OLaPh (Optimal Language Phonemizer) is a multilingual phonemization framework that converts text into phonemes surpassing the quality of …☆23Aug 19, 2026Updated 2 weeks ago
- ☆14Jun 16, 2023Updated 3 years ago
- DETR tensor去除推理过程无用辅助头+fp16部署再次加速+解决转tensorrt 输出全为0问题的新方法。☆12Jan 9, 2024Updated 2 years ago
- ☆11Jun 17, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 机器人人工智能,优达学城cs373作业。 Artificial Intelligence for Robotics, this repository contains all the homework…☆12Nov 12, 2017Updated 8 years ago
- [ICASSP 2026]Official code for "Prosody-Guided Harmonic Attention for Phase-Coherent Neural Vocoding in the Complex Spectrum"☆27Jan 22, 2026Updated 7 months ago
- Convert @NVlabs StyleGAN pkls to @taki0112 StyleGAN-Tensorflow checkpoints (copy over the weights)☆27Sep 17, 2019Updated 6 years ago
- Bongo cat in python! Good for streaming.☆11Jun 6, 2025Updated last year
- 使用OpenCV部署图像描述Image_Captioning,包含C++和Python两个版本的程序☆12Dec 22, 2023Updated 2 years ago
- ☆10May 19, 2022Updated 4 years ago
- 学习用PyTorch创作唐诗☆17Mar 17, 2019Updated 7 years ago