一个用于CosyVoice的api接口项目
☆335Aug 31, 2025Updated last year
Alternatives and similar repositories for cosyvoice-api
Users that are interested in cosyvoice-api are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆33Feb 28, 2025Updated last year
- CosyVoice2 功能扩充(预训练音色推理/3s极速复刻/自然语言控制/自动识别/音色模型保存/API)☆197Mar 13, 2025Updated last year
- 使用vllm加速cosyvoice2的推理☆496Apr 26, 2025Updated last year
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆23,747May 25, 2026Updated 3 months ago
- 基于SenseVoice的funasr版本进行的api发布,可以无缝对接oneapi☆94Aug 12, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 阿里SenseVoice的fastpi封装,采用onnx发布,体积更小,附带量化模型,支持GPU。支持从URL文件进行语音识别。☆114Sep 2, 2024Updated 2 years ago
- 基于官方提供的CosyVoice改造,整体交互适配CosyVoice2模型,开箱即用☆25Jun 15, 2025Updated last year
- CosyVoice在Windows环境下使用的版本☆765Nov 19, 2024Updated last year
- 用于SenseVoice的api项目,输出带时间戳字幕☆49Oct 28, 2024Updated last year
- This repository provides a Docker image for CosyVoice☆27Dec 22, 2024Updated last year
- 🖼️🤖 302 Vector Graphics Generation! 🚀✨☆20Aug 26, 2025Updated last year
- A Bob plugin that calls self-deployed Cosyvoice service to achieve TTS.☆39Aug 13, 2024Updated 2 years ago
- 百聆 是一个类似GPT-4o的语音对话机器人,通过ASR+LLM+TTS实现,集成DeepSeek R1等优秀大模型,接入openClaw,真正的个人语音助手,时延低至800ms,Mac等低配置也可运行,支持打断☆1,775Apr 6, 2026Updated 5 months ago
- This is a speech interaction system built on an open-source model, integrating ASR, LLM, and TTS in sequence. The ASR model is SenceVoice…☆1,276Jun 3, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio…☆9,366Updated this week
- Lightning-responsive CosyVoice streaming API based on FastAPI.☆28Aug 17, 2026Updated last month
- Step-Audio-TTS-3B demo☆14Feb 25, 2025Updated last year
- Added vLLM support to IndexTTS for faster inference.☆1,239Apr 13, 2026Updated 5 months ago
- 🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.☆1,421May 21, 2026Updated 4 months ago
- GLM-4-Voice | 端到端中英语音对话模型☆3,237Dec 5, 2024Updated last year
- 实时交互数字人,可自定义形象与音色,支持音色克隆,对话延迟低至3s。Real-time voice interactive digital human, customizable appearance and voice, supporting voice cloning,…☆1,313Dec 18, 2025Updated 9 months ago
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- ASR_LLM_TTS前端项目☆15Dec 3, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Pseudo Streaming SenseVoice with Hotwords☆470Updated this week
- API and websocket server for sensevoice. It has inherited some enhanced features, such as VAD detection, real-time streaming recognition,…☆542Oct 23, 2024Updated last year
- Just a suturing monster project.☆38Nov 21, 2023Updated 2 years ago
- 小智同学测试工具(websocket)☆45Feb 20, 2025Updated last year
- 开箱即用的本地私有化部署语音服务,快速搭建Qwen3ASR/FunASR与Qwen3TTS/CosyVoice后端☆163Jul 6, 2026Updated 2 months ago
- One command to run ChatTTS☆60Jun 6, 2024Updated 2 years ago
- Real time interactive streaming digital human☆9,608Sep 13, 2026Updated last week
- 一个用于F5-TTS的api和webui项目☆64Dec 25, 2024Updated last year
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆20,481Updated this week
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Enterprise VAD (Voice Activity Detection) in C#.NET (.NET 6.0+) with Microsoft.ML.Net, ONNXRuntime and DirectML. The easiest, efficient, …☆10Apr 20, 2025Updated last year
- 一个简单的本地网页界面,使用ChatTTS将文字合成为语音,同时支持对外提供API接口。A simple native web interface that uses ChatTTS to synthesize text into speech, along with su…☆7,661Jun 14, 2026Updated 3 months ago
- 一个超轻量级、可以在移动端实时运行的数字人模型☆2,647Jul 22, 2026Updated 2 months ago
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆13Jul 15, 2024Updated 2 years ago
- 适用于 GPT-SoVITS 的api调用接口☆350Mar 7, 2024Updated 2 years ago
- Streamer-Sales 销冠 —— 卖货主播 LLM 大模型🛒🎁,一个能够根据给定的商品特点从激发用户购买意愿角度出发进行商品解说的卖货主播大模型。🚀⭐内含详 细的数据生成流程❗ 📦另外还集成了 LMDeploy 加速推理🚀、RAG检索增强生成 📚、TTS文…☆3,775Mar 8, 2025Updated last year
- IndexTTS Fine-tuning notebooks☆140Jun 17, 2025Updated last year