Inference Specialization
☆516Jun 25, 2024Updated 2 years ago
Alternatives and similar repositories for GPT-SoVITS-Inference
Users that are interested in GPT-SoVITS-Inference are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 本项目意图在于让使用各类语音合成引擎的方式变得统一,支持多种语音合成引擎适配器,允许直接作为模组使用或启动后端服务☆770Apr 15, 2024Updated 2 years ago
- 【脱离复杂的环境配置和整合包,极简配置推理服务】从GPT-SoVITS项目里面提取出来的,纯粹的推理服务方案。☆325Apr 11, 2024Updated 2 years ago
- 这是一个批量推理工具,对同一段文字进行多次推理,并且支持随机参数,直到筛选出最满意的结果。☆11Aug 19, 2024Updated last year
- A cli tool for split vocal timbre.☆294Jan 17, 2026Updated 6 months ago
- 1 min voice data can also be used to train a good TTS model! (few shot voice cloning)☆60,723Jul 22, 2026Updated 2 weeks ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 基于中文文本情绪分析自动切换参考音频的 GPT-SoVITS 推理 Demo☆108Mar 8, 2024Updated 2 years ago
- A simple VITS HTTP API, developed by extending Moegoe with additional features.☆1,052May 18, 2026Updated 2 months ago
- 适用于 GPT-SoVITS 的api调用接口☆346Mar 7, 2024Updated 2 years ago
- GPT-SoVITS ONNX Inference Engine & Model Converter☆1,707Aug 3, 2026Updated last week
- 一种基于Emotion2Vec的批量音频情感自动标注脚本☆549Mar 7, 2025Updated last year
- GPT-SoVITS 参考音频推理效果批量试听☆51Mar 8, 2024Updated 2 years ago
- 主要写er-nerf从零到一所有部署过程☆44Aug 28, 2024Updated last year
- Make audio books in one click! Let Genshin characters read novels for you!☆29Aug 2, 2024Updated 2 years ago
- vits2 backbone with multilingual-bert☆8,790Aug 3, 2026Updated last week
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- 通过AI实现对话者的识别并进行文段分割,再接入语音合成,自动生成有声小说☆28Apr 3, 2025Updated last year
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆22,677May 25, 2026Updated 2 months ago
- ☆26Mar 13, 2024Updated 2 years ago
- Bert-VITS2 onnx推理版本☆44Apr 24, 2024Updated 2 years ago
- GPT-SoVITS-V2模型,合并了官方的一些PR,包含但不限于:参考音频自动填充,字幕同步,SillyTavern酒馆接入等功能☆208Jan 15, 2025Updated last year
- ☆68Jul 26, 2025Updated last year
- A lightweight tool that efficiently isolates target speaker data from your datasets.☆20Nov 23, 2024Updated last year
- Easily train a good VC model with voice data <= 10 mins!☆37,300Aug 4, 2026Updated last week
- A generative speech model for daily dialogue.☆39,768Apr 10, 2026Updated 4 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- VC Without Retrain!☆130Apr 27, 2024Updated 2 years ago
- VITS with phoneme-level prosody modeling based on MaskGIT☆85Aug 31, 2024Updated last year
- High-quality multi-lingual text-to-speech library by MyShell.ai. Support English, Spanish, French, Chinese, Japanese and Korean.☆17Sep 18, 2024Updated last year
- GAG is a GUI for GPT-SoVITS inference. Just add it to the official integration package and run for a smoother experience.☆245Jun 24, 2025Updated last year
- SOTA Open Source TTS☆32,128Aug 3, 2026Updated last week
- MuseV: Infinite-length and High Fidelity Virtual Human Video Generation with Visual Conditioned Parallel Denoising☆2,846Jun 28, 2024Updated 2 years ago
- Genshin Datasets For SVC/SVS/TTS☆738Jan 11, 2026Updated 7 months ago
- 低成本的简单基于live2d TTS文字转语音和大模型聊天的直播解决方案☆282Jul 4, 2024Updated 2 years ago
- Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio…☆9,047Jul 27, 2026Updated 2 weeks ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- AI Vtuber是一个由 【ChatterBot/ChatGPT/claude/langchain/chatglm/text-gen-webui/闻达/千问/kimi/ollama】 驱动的虚拟主播【Live2D/UE/xuniren】,可以在 【Bilibili/抖 音/…☆4,427Jul 29, 2025Updated last year
- GeneFace++: Generalized and Stable Real-Time 3D Talking Face Generation; Official Code☆1,809Oct 18, 2024Updated last year
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆19,754Updated this week
- MuseTalk: Real-Time High Quality Lip Synchorization with Latent Space Inpainting☆6,332Sep 26, 2025Updated 10 months ago
- 适用于 diffsinger 的多功能工具集☆10Apr 2, 2023Updated 3 years ago
- Speaker embedding for anime speech domain based on ECAPA_TDNN☆21Jun 22, 2025Updated last year
- ☆20Apr 17, 2025Updated last year