这是一款基于FunASR实现的说话人分离的GUI程序
☆166Aug 25, 2026Updated last month
Alternatives and similar repositories for AudioSeparationGUI
Users that are interested in AudioSeparationGUI are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 这是基于FunASR实现的区分说话人语音识别API | This is a speaker-diarization-based speech recognition API implemented using FunASR.☆28Jun 16, 2026Updated 3 months ago
- API and websocket server for sensevoice. It has inherited some enhanced features, such as VAD detection, real-time streaming recognition,…☆542Oct 23, 2024Updated last year
- Bert-vits2转写和标注独立整合Webui,整合阿里FunAsr,必剪Asr以及Whisper大模型☆181Jul 10, 2024Updated 2 years ago
- Streaming ASR and TTS based on FastAPI+ sherpa-onnx☆226Nov 2, 2025Updated 10 months ago
- 希望用代码为 waifus 绘心。☆101Sep 10, 2026Updated 2 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- MediaEval 2020: Music Mood Classification☆18Mar 5, 2021Updated 5 years ago
- 最棒的的ASR后处理热词方案,基于音素编辑距离,实现热词替换。☆69Jun 10, 2026Updated 3 months ago
- Music Mood Classification on the Million Song Dataset☆17Jul 27, 2019Updated 7 years ago
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆20,517Updated this week
- Make audio books in one click! Let Genshin characters read novels for you!☆29Aug 2, 2024Updated 2 years ago
- About 一個用於接收 WeChatPad 訊息推送的 Webhook 服務端,使用 Python 編寫,支援配置熱加載、簽名驗證、重試機制和日誌記錄,適用於需要對接微信訊息的自動化系統。☆20Jun 27, 2025Updated last year
- 小智同学测试工具(websocket)☆45Feb 20, 2025Updated last year
- This is a project focused on Faster Whisper, a streaming speech recognition project.☆18Sep 27, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Python Wrapper of Silero VAD☆63Updated this week
- 一个简单的音频降噪工具,提高web UI界面和api接口☆48Nov 21, 2024Updated last year
- A Bob plugin that calls self-deployed Cosyvoice service to achieve TTS.☆39Aug 13, 2024Updated 2 years ago
- Real time faster whisper gradio☆24Aug 17, 2025Updated last year
- A talking clock in Chinese for esp32 s3 Box with mp3 player and temperature reading☆12May 7, 2023Updated 3 years ago
- 小智机器人服务端☆18Mar 25, 2025Updated last year
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- 小智的视觉对话☆33Apr 25, 2025Updated last year
- Pseudo Streaming SenseVoice with Hotwords☆471Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 帮助视觉障碍者识别物品,障碍物提醒☆102Mar 14, 2025Updated last year
- Script for Aegisub to cut video and voice files | 在Aegisub中用字幕切割视频和音频文件☆36Oct 13, 2024Updated last year
- 本项目旨在将火山引擎TTS官方API封装为标准MCP协议工具,支持在 Cursor、Claude 等平台通过自然语言参数调用,实现一键语音合成。适用于需要自定义音色、参数映射、自动化批量语音生成的开发者场景。☆16May 22, 2025Updated last year
- Example of onnx quantization☆11Feb 8, 2023Updated 3 years ago
- Maya后台批量导出abc缓存文件☆11Sep 17, 2020Updated 6 years ago
- A Repository for Single- and Multi-modal Speaker Verification, Speaker Recognition and Speaker Diarization☆3,157Dec 8, 2025Updated 9 months ago
- Submission to MediaEval 2021 Emotions and Themes in Music challenge. Noisy-student training for music emotion tagging☆11Dec 2, 2021Updated 4 years ago
- 适用于 NAS、路由器、树莓派等轻量级设备的 xiaozhi-esp32 服务端☆40Sep 4, 2026Updated 3 weeks ago
- 基于 FunASR SenseVoice 模型的实时语音识别服务,支持说话人识别、音频降噪、ASR 错误修正等高级功能。☆22Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- LeNet-5 use c achieve☆13Jan 10, 2020Updated 6 years ago
- 基于roop与codeFormer的换脸一体脚本☆21Apr 9, 2025Updated last year
- 基于SenseVoice的funasr版本进行的api发布,可以无缝对接oneapi☆94Aug 12, 2026Updated last month
- Transcribe Offline by openresearchtools.com is an open source desktop application that allows you to transcribe audio and video fully off…☆18Updated this week
- 📣 商用级开源语音自动识别程序库,开箱即用,全平台支持,中英文混合识别。A Cross-platform implementation of ASR inference. It's based on ONNXRuntime and FunASR. We provide …☆612May 15, 2024Updated 2 years ago
- zero-shot voice conversion & singing voice conversion, with real-time support☆12Feb 11, 2025Updated last year
- 基于Streaming-SenseVoice项目的伪流式实时字幕界面☆14Apr 15, 2025Updated last year