用于SenseVoice的api项目,输出带时间戳字幕
☆49Oct 28, 2024Updated last year
Alternatives and similar repositories for sense-api
Users that are interested in sense-api are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- MT3:多任务多音轨音乐转录的 Gradio 演示。(全中文汉化)☆12Mar 24, 2025Updated last year
- 用于kokoro TTS的webui界面和兼容openai api☆42Feb 4, 2025Updated last year
- speech to text gui for different (e.g. Whisper, Voxtral) models and backends, including whisper.cpp, crispasar, mlx-whisper, faster-whisp…☆31Sep 13, 2026Updated 2 weeks ago
- 一个用于CosyVoice的api接口项目☆335Aug 31, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- 基于Dolphin模型的东方语言音视频转字幕api及webui☆19Apr 3, 2025Updated last year
- local whisper input by Whisper or SenseVoice/FunASR☆22Mar 5, 2025Updated last year
- 一个简单的音频降噪工具,提高web UI界面和api接口☆48Nov 21, 2024Updated last year
- chrome、edge 浏览器插件:视频下载助手☆11Sep 30, 2024Updated last year
- Official implementation of the paper titled "Age and Gender Recognition Using a Convolutional Neural Network with a Specially Designed Mu…☆28Mar 5, 2024Updated 2 years ago
- noise reduction☆17Jul 3, 2024Updated 2 years ago
- 通过LLM进行进行字幕断句分割,处理和优化字幕文件,将自动语音识别(ASR)数据的分段合并与拆分,☆145Dec 17, 2024Updated last year
- 一个用于F5-TTS的api和webui项目☆64Dec 25, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- An N-gram punctuator for Chinese and English.☆20Oct 14, 2025Updated 11 months ago
- semantic tokenizer for speech and music☆20Jul 6, 2025Updated last year
- Official implementation: "AudioSet-R: A Refined AudioSet with Multi-Stage LLM Label Reannotation"☆20Oct 9, 2025Updated 11 months ago
- 简单实现VAD+声纹锁+SenseVoice完成类语音实时转录的小项目☆42Sep 23, 2024Updated 2 years ago
- Streaming ASR and TTS based on FastAPI+ sherpa-onnx☆226Nov 2, 2025Updated 10 months ago
- MusicPlayerNode☆21Feb 25, 2026Updated 7 months ago
- 基于官方提供的CosyVoice改造,整体交互适配CosyVoice2模型,开箱即用☆25Jun 15, 2025Updated last year
- 一个中文语音转文字项目,封装自FireRedASR☆91Feb 24, 2025Updated last year
- TTS-Wrapper makes it easier to use text-to-speech APIs by providing a unified and easy-to-use interface.☆20Jul 26, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ComfyUI implementation of FlashFace: Human Image Personalization with High-fidelity Identity Preservation☆25Jul 31, 2024Updated 2 years ago
- 使用必剪API的语音字幕识别☆420Nov 16, 2025Updated 10 months ago
- Adds alert blockquote support to VS Code's built-in markdown preview☆13Dec 2, 2023Updated 2 years ago
- An open-source project that uses cutting-edge NLP models and real-time web search to provide dynamic voice query responses. Features incl…☆20May 24, 2024Updated 2 years ago
- 基于SenseVoice的funasr版本进行的api发布,可以无缝对接oneapi☆94Aug 12, 2026Updated last month
- Scaled diffusion transformer for text-to-speech synthesis (DiT + T5Gemma2 conditioning, TorchTitan & Megatron backends, tested up to 1024…☆25Mar 29, 2026Updated 5 months ago
- Enhanced CosyVoice with one-click Windows installer, voice management WebUI, and a vLLM-accelerated OpenAI TTS API.☆26Aug 3, 2025Updated last year
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- faster inference☆27Jan 20, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- A lightweight demo of FunASR-Nano using ONNX runtime.☆87Feb 25, 2026Updated 7 months ago
- 开箱即用的本地私有化部署语音服务,快速搭建Qwen3ASR/FunASR与Qwen3TTS/CosyVoice后端☆164Jul 6, 2026Updated 2 months ago
- ☆40Feb 28, 2024Updated 2 years ago
- 深圳大学操作系统作业——制作一个简单的文件管理系统☆24Jul 15, 2020Updated 6 years ago
- ☆35Sep 6, 2025Updated last year
- 一个语音识别项目☆52May 13, 2025Updated last year
- Text-To-Speech for NotebookLM☆39Jul 20, 2025Updated last year