一个用于CosyVoice的api接口项目
☆335Aug 31, 2025Updated 11 months ago
Alternatives and similar repositories for cosyvoice-api
Users that are interested in cosyvoice-api are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆33Feb 28, 2025Updated last year
- CosyVoice2 功能扩充(预训练音色推理/3s极速复刻/自然语言控制/自动识别/音色模型保存/API)☆195Mar 13, 2025Updated last year
- 使用vllm加速cosyvoice2的推理☆498Apr 26, 2025Updated last year
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆22,754May 25, 2026Updated 2 months ago
- 基于SenseVoice的funasr版本进行的api发布,可以无缝对接oneapi☆92Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 阿里SenseVoice的fastpi封装,采用onnx发布,体积更小,附带量化模型,支持GPU。支持从URL文件进行语音识别。☆112Sep 2, 2024Updated last year
- 基于官方提供的CosyVoice改造,整体交互适配CosyVoice2模型,开箱即用☆24Jun 15, 2025Updated last year
- CosyVoice在Windows环境下使用的版本☆768Nov 19, 2024Updated last year
- 用于SenseVoice的api项目,输出带时间戳字幕☆49Oct 28, 2024Updated last year
- This repository provides a Docker image for CosyVoice☆27Dec 22, 2024Updated last year
- 内容审核及速率限制服务☆26May 18, 2025Updated last year
- 🖼️🤖 302 Vector Graphics Generation! 🚀✨☆19Aug 26, 2025Updated 11 months ago
- A Bob plugin that calls self-deployed Cosyvoice service to achieve TTS.☆39Aug 13, 2024Updated 2 years ago
- 百聆 是一个类似GPT-4o的语音对话机器人,通过ASR+LLM+TTS实现,集成DeepSeek R1等优秀大模型,接入openClaw,真正的个人语音助手,时延低至800ms,Mac等低配置也可运行,支持打断☆1,753Apr 6, 2026Updated 4 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- This is a speech interaction system built on an open-source model, integrating ASR, LLM, and TTS in sequence. The ASR model is SenceVoice…☆1,267Jun 3, 2026Updated 2 months ago
- Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio…☆9,071Updated this week
- Lightning-responsive CosyVoice streaming API based on FastAPI.☆28Updated this week
- Step-Audio-TTS-3B demo☆14Feb 25, 2025Updated last year
- Added vLLM support to IndexTTS for faster inference.☆1,225Apr 13, 2026Updated 4 months ago
- 🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.☆1,414May 21, 2026Updated 2 months ago
- GLM-4-Voice | 端到端中英语音对话模型☆3,215Dec 5, 2024Updated last year
- 实时交互数字人,可自定义形象与音色,支持音色克隆,对话延迟低至3s。Real-time voice interactive digital human, customizable appearance and voice, supporting voice cloning,…☆1,301Dec 18, 2025Updated 7 months ago
- ASR_LLM_TTS前端项目☆15Dec 3, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- Pseudo Streaming SenseVoice with Hotwords☆469Jun 15, 2026Updated 2 months ago
- API and websocket server for sensevoice. It has inherited some enhanced features, such as VAD detection, real-time streaming recognition,…☆539Oct 23, 2024Updated last year
- Just a suturing monster project.☆38Nov 21, 2023Updated 2 years ago
- 小智同学测试工具(websocket)☆45Feb 20, 2025Updated last year
- 开箱即用的本地私有化部署语音服务,快速搭建Qwen3ASR/FunASR与Qwen3TTS/CosyVoice后端☆155Jul 6, 2026Updated last month
- One command to run ChatTTS☆60Jun 6, 2024Updated 2 years ago
- Real time interactive streaming digital human☆8,754Aug 7, 2026Updated last week
- 一个用于F5-TTS的api和webui项目☆63Dec 25, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆19,820Updated this week
- Enterprise VAD (Voice Activity Detection) in C#.NET (.NET 6.0+) with Microsoft.ML.Net, ONNXRuntime and DirectML. The easiest, efficient, …☆10Apr 20, 2025Updated last year
- 一个简单的本地网页界面,使用ChatTTS将文字合成为语音,同时支持对外提供API接口。A simple native web interface that uses ChatTTS to synthesize text into speech, along with su…☆7,633Jun 14, 2026Updated 2 months ago
- 一个超轻量级、可以在移动端实时运行的数字人模型☆2,620Jul 22, 2026Updated 3 weeks ago
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆13Jul 15, 2024Updated 2 years ago
- 适用于 GPT-SoVITS 的api调用接口☆346Mar 7, 2024Updated 2 years ago
- Streamer-Sales 销冠 —— 卖货主播 LLM 大模型🛒🎁,一个能够根据给定的商品特点从激发用户购买意愿角度出发进行商品解说的卖货主播大模型。🚀⭐内含详细的数据生成流程❗ 📦另外还集成了 LMDeploy 加速推理🚀、RAG检索增强生成 📚、TTS文…☆3,753Mar 8, 2025Updated last year