简单实现VAD+声纹锁+SenseVoice完成类语音实时转录的小项目
☆42Sep 23, 2024Updated last year
Alternatives and similar repositories for SenseVoice-Real-Time
Users that are interested in SenseVoice-Real-Time are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- Pseudo Streaming SenseVoice with Hotwords☆469Jun 15, 2026Updated last month
- 基于Fastrtc、Ollama、FunASR和MegaTTS的大模型中文语音实时对话应用☆22Apr 26, 2025Updated last year
- This is a project focused on Faster Whisper, a streaming speech recognition project.☆18Sep 27, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- local whisper input by Whisper or SenseVoice/FunASR☆22Mar 5, 2025Updated last year
- 使用fastrtc框架调用qwen-2.5-omni-realtime实现实时语音、视频等☆14Jun 27, 2025Updated last year
- 这是基于FunASR实现的区分说话人语音识别API | This is a speaker-diarization-based speech recognition API implemented using FunASR.☆28Jun 16, 2026Updated last month
- ☆18Nov 27, 2024Updated last year
- This repository contains Nanny-On agent source codes. Anyone can contribute to write and update the codes. Please join us.☆17Nov 22, 2021Updated 4 years ago
- API and websocket server for sensevoice. It has inherited some enhanced features, such as VAD detection, real-time streaming recognition,…☆539Oct 23, 2024Updated last year
- SenseVoice-python: A enterprise-grade open source multi-language asr system from funasr opensource with onnxruntime☆114Jun 12, 2026Updated 2 months ago
- CTC decoder with hotwords for ASR.☆39Aug 5, 2026Updated last week
- A enterprise-grade Chinese-English code switch punctuator from funasr.☆34Apr 26, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Data manipulation and transformation for audio signal processing, powered by PyTorch☆10Sep 30, 2024Updated last year
- Python Wrapper for RnNoise v0.2☆78Jan 14, 2026Updated 6 months ago
- paraformer(chinense asr) online onnx runtime for python☆54Mar 27, 2024Updated 2 years ago
- An ASR API server for FunASR☆61Jul 18, 2026Updated 3 weeks ago
- paraformer web server build with sanic☆28May 3, 2023Updated 3 years ago
- ☆52Nov 26, 2023Updated 2 years ago
- This is a project based on machine learning and deep learning method for playing Gobang by controlling mechanical arm(利用机械臂下五子棋)☆13Apr 16, 2023Updated 3 years ago
- Unity AudioDance☆10Sep 12, 2020Updated 5 years ago
- ☆20Feb 9, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- One command to start a streaming ASR server.☆12Oct 2, 2024Updated last year
- 用于SenseVoice的api项目,输出带时间戳字幕☆49Oct 28, 2024Updated last year
- Port of Funasr's Paraformer model in C/C++☆42Jun 19, 2024Updated 2 years ago
- tts fronted-end☆11Dec 19, 2018Updated 7 years ago
- 通过 Python 驱动 ADB , 实现自动刷抖音,并将无水印视频下载至本地.☆13Jul 9, 2021Updated 5 years ago
- ☆12Jul 11, 2024Updated 2 years ago
- ASR (Automatic Speech Recognition) for real-time streamed audio powered by Whisper and tranformers☆36Apr 22, 2026Updated 3 months ago
- 基于SenseVoice的funasr版本进行的api发布,可以无缝对接oneapi☆92Updated this week
- Code to reproduce paper results (or as close as possible, depending on data-availability). Each publication has a Jupyter notebook. Mostl…☆12Mar 8, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Superresolution running on Rockchip NPU (RK3588, etc..)☆24Jul 7, 2024Updated 2 years ago
- MP3, MP4(m4a, m4b), FLAC and OGG(Vorbis, Opus) meta data reader and writer for go☆11Dec 29, 2024Updated last year
- ☆19Jul 19, 2025Updated last year
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- Simple optimized UI effect. Inspired by Marerial Design☆16Apr 26, 2023Updated 3 years ago
- 基于FastAPI的语音服务系统,集成语音合成(TTS)和语音识别(STT)功能。使用CosyVoice2作为TTS引擎,FunASR作为STT引擎,支持零样本语音克隆、流式输出、多种语言识别等高级功能。☆21Apr 1, 2025Updated last year
- A simple implementation for improving CosyVoice2 by GRPO method☆39May 5, 2026Updated 3 months ago