用于SenseVoice的api项目,输出带时间戳字幕
☆49Oct 28, 2024Updated last year
Alternatives and similar repositories for sense-api
Users that are interested in sense-api are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 用于kokoro TTS的webui界面和兼容openai api☆41Feb 4, 2025Updated last year
- MT3:多任务多音轨音乐转录的 Gradio 演示。(全中文汉化)☆12Mar 24, 2025Updated last year
- speech to text gui for different (e.g. Whisper, Voxtral) models and backends, including whisper.cpp, crispasar, mlx-whisper, faster-whisp…☆29Aug 16, 2026Updated 2 weeks ago
- 一个用于CosyVoice的api接口项目☆335Aug 31, 2025Updated last year
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- local whisper input by Whisper or SenseVoice/FunASR☆22Mar 5, 2025Updated last year
- 基于Dolphin模型的东方语言音视频转字幕api及webui☆19Apr 3, 2025Updated last year
- 一个简单的音频降噪工具,提高web UI界面和api接口☆46Nov 21, 2024Updated last year
- chrome、edge 浏览器插件:视频下载助手☆11Sep 30, 2024Updated last year
- Official implementation of the paper titled "Age and Gender Recognition Using a Convolutional Neural Network with a Specially Designed Mu…☆28Mar 5, 2024Updated 2 years ago
- noise reduction☆17Jul 3, 2024Updated 2 years ago
- 通过LLM进行进行字幕断句分割,处理和优化字幕文件,将自动语音识别(ASR)数据的分段合并与拆分,☆144Dec 17, 2024Updated last year
- An N-gram punctuator for Chinese and English.☆20Oct 14, 2025Updated 10 months ago
- semantic tokenizer for speech and music☆20Jul 6, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆23Oct 17, 2024Updated last year
- Official implementation: "AudioSet-R: A Refined AudioSet with Multi-Stage LLM Label Reannotation"☆20Oct 9, 2025Updated 10 months ago
- 简单实现VAD+声纹锁+SenseVoice完成类语音实时转录的小项目☆42Sep 23, 2024Updated last year
- Streaming ASR and TTS based on FastAPI+ sherpa-onnx☆225Nov 2, 2025Updated 10 months ago
- 基于ddtv开放的api开发的一套webUI,基于VUE 3.0开发。☆12Apr 3, 2023Updated 3 years ago
- 基于官方提供的CosyVoice改造,整体交互适配CosyVoice2模型,开箱即用☆24Jun 15, 2025Updated last year
- 一个中文语音转文字项目,封装自FireRedASR☆88Feb 24, 2025Updated last year
- TTS-Wrapper makes it easier to use text-to-speech APIs by providing a unified and easy-to-use interface.☆20Jul 26, 2024Updated 2 years ago
- ComfyUI implementation of FlashFace: Human Image Personalization with High-fidelity Identity Preservation☆25Jul 31, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 使用必剪API的语音字幕识别☆418Nov 16, 2025Updated 9 months ago
- ICO Source Spider, write in NodeJS☆12May 4, 2018Updated 8 years ago
- 第三方 录播姬/BLREC/等 录播机管理面板☆22Jun 27, 2026Updated 2 months ago
- 基于SenseVoice的funasr版本进行的api发布,可以无缝对接oneapi☆94Aug 12, 2026Updated 3 weeks ago
- Scaled diffusion transformer for text-to-speech synthesis (DiT + T5Gemma2 conditioning, TorchTitan & Megatron backends, tested up to 1024…☆24Mar 29, 2026Updated 5 months ago
- Enhanced CosyVoice with one-click Windows installer, voice management WebUI, and a vLLM-accelerated OpenAI TTS API.☆24Aug 3, 2025Updated last year
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- faster inference☆27Jan 20, 2025Updated last year
- A test web browser using rust and egui☆13Apr 13, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- BILIBILI COOKIE 管理器☆18Jul 20, 2026Updated last month
- A lightweight demo of FunASR-Nano using ONNX runtime.☆87Feb 25, 2026Updated 6 months ago
- 开箱即用的本地私有化部署语音服务,快速搭建Qwen3ASR/FunASR与Qwen3TTS/CosyVoice后端☆160Jul 6, 2026Updated last month
- ☆40Feb 28, 2024Updated 2 years ago
- Omni SenseVoice: High-Speed Speech Recognition with words timestamps 🗣️🎯☆897Dec 10, 2025Updated 8 months ago
- ☆35Sep 6, 2025Updated 11 months ago
- (Experimental) Predicting hand assignments in piano MIDI using neural networks☆13Oct 11, 2024Updated last year