将音频或视频中的中文语音识别并导出为srt字幕,基于魔塔社区Paraformer模型
☆116Jul 10, 2024Updated 2 years ago
Alternatives and similar repositories for zh_recogn
Users that are interested in zh_recogn are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 基于ffmpeg.wasm的在线视频处理工具☆57May 15, 2024Updated 2 years ago
- 张艺谋(国师)一键声音克隆和恶搞文本生成项目☆17Jun 15, 2023Updated 3 years ago
- Finding the most similar tone/color in a large collection of audio. 在一大堆音频中寻找最相似的音色。☆13Jun 17, 2024Updated 2 years ago
- 基于 RWKV_Role_Playing 项目接入GPT-SoVITS语音对话项目☆30Apr 8, 2024Updated 2 years ago
- Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式☆4,749Jan 22, 2026Updated 6 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Api tool for local offline text translation supporting multiple languages/支持多语言的本地离线文字翻译api☆482Nov 4, 2024Updated last year
- Bert-vits2转写和标注独立整合Webui,整合阿里FunAsr,必剪Asr以及Whisper大模型☆182Jul 10, 2024Updated 2 years ago
- 基于Python3.10异步非阻塞框架Tornado6.0和前端Vue.js3框架实现ChatGPT的流式返回协议Server-sent events☆23Mar 7, 2023Updated 3 years ago
- 数据集自动化制作脚本☆71Mar 26, 2023Updated 3 years ago
- ChatTTS 2000条音色稳定性打分🥇+区分男女年龄👧+在线试听🔈 ChatTTS 2K Speaker Stability Score & Categorized by Gender and Age & Audio Preview☆732Jul 2, 2024Updated 2 years ago
- A sound cloning tool with a web interface, using your voice or any sound to record audio / 一个带web界面的声音克隆工具,使用你的音色或任意声音来录制音频☆8,986Aug 29, 2025Updated 11 months ago
- 基于Bert-vits2-Extra项目添加的流式推理和流式接口api功能☆16Apr 12, 2024Updated 2 years ago
- 文本语料转训练集工具,txt转dataset☆91May 1, 2024Updated 2 years ago
- an extremely simple tool for separating vocals and background music, completely localized for web operation, using 2stems/4stems/5stems …☆1,999Nov 26, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 利用Github Action的能力获取天气并生成图片,用于第三方分享☆15Mar 17, 2026Updated 5 months ago
- IMAGdressing在Windows环境下运行的webui界面☆21Jul 25, 2024Updated 2 years ago
- Direct Preference Optimization for RWKV, aiming for RWKV-5 and 6.☆11Mar 1, 2024Updated 2 years ago
- 音频响度统一,音量归一化处理☆13May 3, 2024Updated 2 years ago
- Translate the video from one language to another and embed dubbing & subtitles.☆18,708Updated this week
- 复刻自TG群俊于phper的库:https://github.com/q215613905/TVBoxOS/☆16Updated this week
- FitGamerAI 是一个基于 Next.js 的创新体感互动游戏平台。☆17Mar 9, 2026Updated 5 months ago
- 可循环值守和多人录制的直播录制软件,支持抖音、TikTok、快手、虎牙、斗鱼、B站、小红书、网易cc、pandaTV等平台直播录制,抓取多平台直播源地址,抖音无水印解析,快手无水印解析☆18Feb 8, 2024Updated 2 years ago
- Python的音频工具☆16Dec 5, 2025Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Sequence alignement methods with helpers for PyTorch.☆24Nov 30, 2022Updated 3 years ago
- SenseVoice-python: A enterprise-grade open source multi-language asr system from funasr opensource with onnxruntime☆114Jun 12, 2026Updated 2 months ago
- 在cloudflare上基于m2m100创建完全免费的翻译API服务☆56Nov 8, 2024Updated last year
- 使用深度学习框架提取视频硬字幕;docker容器免安装深度学习库,使用本地api接口使得界面和后端识别分离;☆23Dec 20, 2021Updated 4 years ago
- llama_index_examples_python3.10 基于ChatGPT的垂直领域语料向量索引优化☆36Apr 10, 2023Updated 3 years ago
- 跨语种语音克隆,中文版Webui☆63Jan 4, 2024Updated 2 years ago
- 启发自Nuxt3的前端模板库☆12Mar 16, 2022Updated 4 years ago
- 根据声音生成音色文件☆35Aug 6, 2024Updated 2 years ago
- Wechat robot for Java 使用Java开发的微信机器人程序☆18Sep 12, 2016Updated 9 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- 一个极简的音视频格式转换工具☆22Mar 12, 2024Updated 2 years ago
- 一个有想法的视频处理工具,追求 AI 效果☆59Mar 9, 2022Updated 4 years ago
- Reverse proxies for PHP☆11Aug 16, 2024Updated 2 years ago
- Unsloth框架在Windows平台微调训练Qwen2大模型,非WSL☆61Jun 19, 2024Updated 2 years ago
- AI powered speech denoising and enhancement. Adapted for windows and optimized☆93Jul 12, 2024Updated 2 years ago
- 一个文本转语音工具,可以帮你把生硬的文字变成灵动的语音。☆51Nov 9, 2024Updated last year
- 基于Read-aloud项目增加GPT-SoVits接口,Chrome扩展,大声朗读☆25Feb 27, 2024Updated 2 years ago