This is a speech interaction system built on an open-source model, integrating ASR, LLM, and TTS in sequence. The ASR model is SenceVoice, the LLM models are QWen2.5-0.5B/1.5B, and there are three TTS models: CosyVoice, Edge-TTS, and pyttsx3
☆1,262Jun 3, 2026Updated last month
Alternatives and similar repositories for ASR-LLM-TTS
Users that are interested in ASR-LLM-TTS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 百聆 是一个类似GPT-4o的语音对话机器人,通过ASR+LLM+TTS实现,集成DeepSeek R1等优秀大模型,接入openClaw,真正的个人语音助手,时延低至800ms,Mac等低配置也可运行,支持打断☆1,742Apr 6, 2026Updated 3 months ago
- API and websocket server for sensevoice. It has inherited some enhanced features, such as VAD detection, real-time streaming recognition,…☆538Oct 23, 2024Updated last year
- Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio…☆8,935Updated this week
- 实时交互数字人,可自定义形象与音色,支持音色克隆,对话延迟低至3s。Real-time voice interactive digital human, customizable appearance and voice, supporting voice cloning,…☆1,296Dec 18, 2025Updated 7 months ago
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆22,400May 25, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆19,467Updated this week
- This is a multi-character, ultra-personalized StoryTeller. It includes: 1) efficiently and accurately build multi-character voice library …☆64Feb 2, 2025Updated last year
- This is a web-based intelligent dialogue program built using ASR, LLM, and TTS.☆26Dec 3, 2024Updated last year
- 实时STT,连接OpenAI接口/智谱AI(流式LLM)和GPT-SOVITS/Edge-TTS,通过网页的方式,进行跨网络的服务调用,实现实时对话的效果☆433Dec 31, 2024Updated last year
- 1688找供应商 —— 结合用户需求与关键字查询对应的供应商及工厂信息 核心工具能力:1688供应商查询能力。用于查询1688平台上的供应商及工厂信息。 触发词:找供应商、查供应商、1688供应商、供应商信息、工厂信息、产业带查询。 不触发场景:找商品/选品 → 1688-…☆551May 7, 2026Updated 2 months ago
- 88生意通是1688线下B2B交易的得力帮手,一句话搞定全流程操作!无论您是卖家还是买家,只需一句指令,即可轻松完成交易单创建、签署、确认收货、退款等核心操作,全面支持账号状态查询、实名认证、绑卡及交易,让每一步交易流程更清晰、更可控。通过智能化交互,实现交易流程数字化,提…☆575Mar 27, 2026Updated 3 months ago
- 本项目为xiaozhi-esp32提供后端服务,帮助您快速搭建ESP32设备控制服务器。Backend service for xiaozhi-esp32, helps you quickly build an ESP32 device control server.☆10,133Updated this week
- ASR_LLM_TTS前端项目☆15Dec 3, 2024Updated last year
- 一个结合了ASR+LLM+TTS+监控的多功能AI机器人。支持所有以open ai为API调用格式的模型。支持LLM模型流式输出,以及对话打断、视频对话☆30Apr 15, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 1688 商家重点品圈选 —— 基于多维度商品评分智能识别值得重点运营的商品或搜索商品。 工具能力:五维度评分(销售贡献、流量效率、成长潜力、营销ROI、商品健康度),商品分层(S/A/B/C级),搜索商品。 触 发词:重点品查看、圈选重点品、圈选运营商品、今日运营重点、选品…☆623May 11, 2026Updated 2 months ago
- A Repository for Single- and Multi-modal Speaker Verification, Speaker Recognition and Speaker Diarization☆3,070Dec 8, 2025Updated 7 months ago
- 一个用于CosyVoice的api接口项目☆333Aug 31, 2025Updated 10 months ago
- Streaming ASR and TTS based on FastAPI+ sherpa-onnx☆222Nov 2, 2025Updated 8 months ago
- ☆19Jul 19, 2025Updated last year
- Pseudo Streaming SenseVoice with Hotwords☆467Jun 15, 2026Updated last month
- Real time interactive streaming digital human☆8,499Jul 19, 2026Updated last week
- Open-source AI assistant ecosystem with MCP integrations, multimodal workflows, IoT support, and cross-platform voice interaction.☆3,423Updated this week
- Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime…☆13,783Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Lightning-responsive CosyVoice streaming API based on FastAPI.☆28Apr 27, 2026Updated 2 months ago
- A demo project for idempotent pipeline design, duplicate prevention, and retry-safe processing.☆232Mar 20, 2026Updated 4 months ago
- 使用vllm加速cosyvoice2的推理☆498Apr 26, 2025Updated last year
- A free detector capable of identifying content generated by all advanced AI models.☆373Updated this week
- 异步语音对话组件。☆32Mar 13, 2025Updated last year
- GLM-4-Voice | 端到端中英语音对话模型☆3,209Dec 5, 2024Updated last year
- An MCP-based chatbot | 一个基于MCP的聊天机器人☆28,351Updated this week
- Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR be…☆1,940Feb 25, 2026Updated 5 months ago
- ☆33Feb 28, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 简单实现VAD+声纹锁+SenseVoice完成类语音实时转录的小项目☆42Sep 23, 2024Updated last year
- Sample Repository for the AlibabaCloud Bailian Speech SDK☆421Jul 14, 2026Updated last week
- Build your own AI friend☆778Jun 7, 2025Updated last year
- Real-time causal inference framework for Web Vitals optimization using streaming SCMs, public CrUX/PageSpeed field data, and auditable in…☆400Jul 2, 2026Updated 3 weeks ago
- ☆3,646Jun 9, 2026Updated last month
- 一个模块化,全过程可离线,低占用率的对话机器人/智能音箱☆159Mar 25, 2026Updated 4 months ago
- 财税税负测算专家技能。帮助用户计算商品含税定价、评估整体税负。支持增值税测算、四税联算(增值税+附加税+印花税+所得税)、整体税负分析等场景。用户提到税务、税率、含税定价、税负、增值税、所得税、发票、纳税人等测算问题时使用。☆696Apr 23, 2026Updated 3 months ago