一个用于CosyVoice的api接口项目
☆335Aug 31, 2025Updated last year
Alternatives and similar repositories for cosyvoice-api
Users that are interested in cosyvoice-api are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆33Feb 28, 2025Updated last year
- CosyVoice2 功能扩充(预训练音色推理/3s极速复刻/自然语言控制/自动识别/音色模型保存/API)☆197Mar 13, 2025Updated last year
- 使用vllm加速cosyvoice2的推理☆497Apr 26, 2025Updated last year
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆23,425May 25, 2026Updated 3 months ago
- 基于SenseVoice的funasr版本进行的api发布,可以无缝对接oneapi☆94Aug 12, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 阿里SenseVoice的fastpi封装,采用onnx发布,体积更小,附带量化模型,支持GPU。支持从URL文件进行语音识别。☆113Sep 2, 2024Updated 2 years ago
- 基于官方提供的CosyVoice改造,整体交互适配CosyVoice2模型,开箱即用☆24Jun 15, 2025Updated last year
- CosyVoice在Windows环境下使用的版本☆767Nov 19, 2024Updated last year
- 用于SenseVoice的api项目,输出带时间戳字幕☆49Oct 28, 2024Updated last year
- This repository provides a Docker image for CosyVoice☆27Dec 22, 2024Updated last year
- 🖼️🤖 302 Vector Graphics Generation! 🚀✨☆20Aug 26, 2025Updated last year
- A Bob plugin that calls self-deployed Cosyvoice service to achieve TTS.☆39Aug 13, 2024Updated 2 years ago
- 百聆 是一个类似GPT-4o的语音对话机器人,通过ASR+LLM+TTS实现,集成DeepSeek R1等优秀大模型,接入openClaw,真正的个人语音助手,时延低至800ms,Mac等低配置也可运行,支持打断☆1,759Apr 6, 2026Updated 4 months ago
- This is a speech interaction system built on an open-source model, integrating ASR, LLM, and TTS in sequence. The ASR model is SenceVoice…☆1,275Jun 3, 2026Updated 3 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio…☆9,219Updated this week
- Lightning-responsive CosyVoice streaming API based on FastAPI.☆28Aug 17, 2026Updated 2 weeks ago
- Step-Audio-TTS-3B demo☆14Feb 25, 2025Updated last year
- Added vLLM support to IndexTTS for faster inference.☆1,236Apr 13, 2026Updated 4 months ago
- 🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.☆1,418May 21, 2026Updated 3 months ago
- GLM-4-Voice | 端到端中英语音对话模型☆3,229Dec 5, 2024Updated last year
- 实时交互数字人,可自定义形象与音色,支持音色克隆,对话延迟低至3s。Real-time voice interactive digital human, customizable appearance and voice, supporting voice cloning,…☆1,303Dec 18, 2025Updated 8 months ago
- ASR_LLM_TTS前端项目☆15Dec 3, 2024Updated last year
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Pseudo Streaming SenseVoice with Hotwords☆470Jun 15, 2026Updated 2 months ago
- API and websocket server for sensevoice. It has inherited some enhanced features, such as VAD detection, real-time streaming recognition,…☆542Oct 23, 2024Updated last year
- Just a suturing monster project.☆38Nov 21, 2023Updated 2 years ago
- 小智同学测试工具(websocket)☆45Feb 20, 2025Updated last year
- 开箱即用的本地私有化部署语音服务,快速搭建Qwen3ASR/FunASR与Qwen3TTS/CosyVoice后端☆160Jul 6, 2026Updated last month
- One command to run ChatTTS☆60Jun 6, 2024Updated 2 years ago
- Real time interactive streaming digital human☆9,376Updated this week
- 一个用于F5-TTS的api和webui项目☆63Dec 25, 2024Updated last year
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆20,151Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Enterprise VAD (Voice Activity Detection) in C#.NET (.NET 6.0+) with Microsoft.ML.Net, ONNXRuntime and DirectML. The easiest, efficient, …☆10Apr 20, 2025Updated last year
- 一个简单的本地网页界面,使用ChatTTS将文字合成为语音,同时支持对外提供API接口。A simple native web interface that uses ChatTTS to synthesize text into speech, along with su…☆7,650Jun 14, 2026Updated 2 months ago
- 一个超轻量级、可以在移动端实时运行的数字人模型☆2,635Jul 22, 2026Updated last month
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆13Jul 15, 2024Updated 2 years ago
- 适用于 GPT-SoVITS 的api调用接口☆349Mar 7, 2024Updated 2 years ago
- Streamer-Sales 销冠 —— 卖货主播 LLM 大模型🛒🎁,一个能够根据给定的商品特点从激发用户购买意愿角度出发进行商品解说的 卖货主播大模型。🚀⭐内含详细的数据生成流程❗ 📦另外还集成了 LMDeploy 加速推理🚀、RAG检索增强生成 📚、TTS文…☆3,765Mar 8, 2025Updated last year
- IndexTTS Fine-tuning notebooks☆139Jun 17, 2025Updated last year