这是一款基于FunASR实现的说话人分离的GUI程序
☆166Dec 14, 2025Updated 8 months ago
Alternatives and similar repositories for AudioSeparationGUI
Users that are interested in AudioSeparationGUI are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is a web-based intelligent dialogue program built using ASR, LLM, and TTS.☆26Dec 3, 2024Updated last year
- API and websocket server for sensevoice. It has inherited some enhanced features, such as VAD detection, real-time streaming recognition,…☆539Oct 23, 2024Updated last year
- Streaming ASR and TTS based on FastAPI+ sherpa-onnx☆225Nov 2, 2025Updated 9 months ago
- 希望用代码为 waifus 绘心。☆101Jul 28, 2026Updated 2 weeks ago
- 最棒的的ASR后处理热词方案,基于音素编辑距离,实现热词替换。☆50Jun 10, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Music Mood Classification on the Million Song Dataset☆17Jul 27, 2019Updated 7 years ago
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆19,865Updated this week
- Make audio books in one click! Let Genshin characters read novels for you!☆29Aug 2, 2024Updated 2 years ago
- 快速提取音视频内容,整理成一份结构化的markdown笔记☆2,262Updated this week
- This is a project focused on Faster Whisper, a streaming speech recognition project.☆18Sep 27, 2024Updated last year
- 小智同学测试工具(websocket)☆45Feb 20, 2025Updated last year
- Python Wrapper of Silero VAD☆63May 8, 2025Updated last year
- 一个简单的音频降噪工具,提高web UI界面和api接口☆46Nov 21, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A Bob plugin that calls self-deployed Cosyvoice service to achieve TTS.☆39Aug 13, 2024Updated 2 years ago
- C++ version of pyannote audio overlapped speech detection pipeline☆13Feb 14, 2024Updated 2 years ago
- 小智机器人服务端☆18Mar 25, 2025Updated last year
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- 小智的视觉对话☆33Apr 25, 2025Updated last year
- Pseudo Streaming SenseVoice with Hotwords☆469Jun 15, 2026Updated 2 months ago
- 帮助视觉障碍者识别物品,障碍物提醒☆100Mar 14, 2025Updated last year
- Script for Aegisub to cut video and voice files | 在Aegisub中用字幕切割视频和音频文件☆35Oct 13, 2024Updated last year
- 本项目旨在将火山引擎TTS官方API封装为标准MCP协议工具,支持在 Cursor、Claude 等平台通过自然语言参数调用,实现一键语音合成。适用于需要自定义音色、参数映射、自动化批量语音生成的开发者场景。☆16May 22, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A enterprise-grade Chinese-English code switch punctuator from funasr.☆34Apr 26, 2024Updated 2 years ago
- Submission to MediaEval 2021 Emotions and Themes in Music challenge. Noisy-student training for music emotion tagging☆11Dec 2, 2021Updated 4 years ago
- 适用于 NAS、路由器、树莓派等轻量级设备的 xiaozhi-esp32 服务端☆41May 7, 2026Updated 3 months ago
- 基于 FunASR SenseVoice 模型的实时语音识别服务,支持说话人识别、音频降噪、ASR 错误修正等高级功能。☆21Jul 10, 2026Updated last month
- LeNet-5 use c achieve☆13Jan 10, 2020Updated 6 years ago
- 基于SenseVoice的funasr版本进行的api发布,可以无缝对接oneapi☆94Updated this week
- 📣 商用级开源语音自动识别程序库,开箱即用,全平台支持,中英文混合识别。A Cross-platform implementation of ASR inference. It's based on ONNXRuntime and FunASR. We provide …☆611May 15, 2024Updated 2 years ago
- zero-shot voice conversion & singing voice conversion, with real-time support☆11Feb 11, 2025Updated last year
- 基于Streaming-SenseVoice项目的伪流式实时字幕界面☆13Apr 15, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"☆14Nov 17, 2024Updated last year
- 一个实时交互的语音项目☆54May 20, 2026Updated 2 months ago
- ☆10Jan 6, 2025Updated last year
- This is a speech interaction system built on an open-source model, integrating ASR, LLM, and TTS in sequence. The ASR model is SenceVoice…☆1,268Jun 3, 2026Updated 2 months ago
- ComfyUI sampler for HyperSDXL UNet☆11Jun 20, 2024Updated 2 years ago
- ☆20Feb 9, 2026Updated 6 months ago
- Noto fonts component for xiaozhi-esp32☆22Jul 19, 2026Updated 3 weeks ago