stt websockect server using sherpa-onnx
☆61Feb 28, 2026Updated 7 months ago
Alternatives and similar repositories for stt-server
Users that are interested in stt-server are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- On-device speech AI runtime for ASR, TTS, VAD, and voice cloning. Python-simple, C++-native, GGUF-powered.☆29Aug 7, 2026Updated last month
- Port of Funasr's Paraformer model in C/C++☆43Jun 19, 2024Updated 2 years ago
- 一个基于 Sherpa-ONNX 的高性能语音识别服务,支持实时VAD(语音活动检测)、多语言语音识别和声纹识别功能。☆116Jan 4, 2026Updated 8 months ago
- SummerTTS 是一个基于C++的独立编译的中文和英文语音合成项目,可以本地运行不需要网络,而且没有额外的依赖,一键编译完成即可用于中文和英文的语音合成。SummerTTS is a standalone Chinese and English speech synt…☆25Aug 17, 2024Updated 2 years ago
- ☆25Mar 8, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Streaming ASR and TTS based on FastAPI+ sherpa-onnx☆226Nov 2, 2025Updated 10 months ago
- Port of Funasr's Sense-voice model in C/C++☆574Dec 19, 2025Updated 9 months ago
- Moss: A voice assistant using LLM and Langchain which can control your home assistant and chat more.☆34Jul 12, 2025Updated last year
- ☆50Jan 20, 2025Updated last year
- To convert CosyVoice model to ONNX☆16Dec 22, 2025Updated 9 months ago
- FreeSWITCH Module for MP3 recording☆11Feb 23, 2020Updated 6 years ago
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- A lightweight demo of FunASR-Nano using ONNX runtime.☆87Feb 25, 2026Updated 7 months ago
- High-performance Qwen3-TTS implementation | Instruction-driven · Zero-shot voice cloning · Streaming · RTF 0.55☆70Jun 26, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- low-latency realtime ASR based on FireRedASR☆61Jul 8, 2025Updated last year
- ☆13Mar 30, 2023Updated 3 years ago
- Hacked FreeSWITCH-G729 speech codec using Intel® Integrated Performance Primitive.☆23Jan 22, 2013Updated 13 years ago
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆35Sep 25, 2025Updated last year
- August智能体框架的相关功能包,已成功部署在Robonova机器人上☆20Aug 31, 2025Updated last year
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- A playground for experimenting with acoustic echo cancellation using a microphone, speaker, and ONNX.☆13Oct 22, 2024Updated last year
- Freeswitch Speech-To-Text module☆17Aug 25, 2026Updated last month
- A enterprise-grade Voice Activity Detector from modelscope and funasr.☆142Apr 26, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The CPP version of Silero VAD: pre-trained enterprise-grade Voice Activity Detector☆23May 11, 2024Updated 2 years ago
- TEN VAD low-latency voice activity detection for real-time streaming, integrated with livekit-agents☆26Nov 13, 2025Updated 10 months ago
- real-time web visualizer for 3D gaussian splatting☆10Jan 31, 2025Updated last year
- Pseudo Streaming SenseVoice with Hotwords☆471Updated this week
- 一键将视频转换为优质小红书笔记,自动优化内容和配图;追加了可以读取本地视频的功能☆12Dec 22, 2024Updated last year
- A package used to test webrtc apm functions, such as aec, ns☆17Feb 21, 2019Updated 7 years ago
- A streaming audio reader, processor, and writer built on top of soundfile, and PyAV (bindings for FFmpeg)☆39Aug 27, 2026Updated last month
- An open source chat bot architecture for voice/vision (and multimodal) assistants, local(CPU/GPU bound) and remote(I/O bound) to run.☆90Dec 28, 2025Updated 8 months ago
- silero-vad pytorch implement☆38Nov 23, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Clean up noisy speech in real time with DPDFNet - open-source streaming speech enhancement for research, audio apps, and edge devices. In…☆156Updated this week
- Template for creating audio encoders compatible with X-ARES☆19Feb 11, 2026Updated 7 months ago
- paraformer web server build with sanic☆29May 3, 2023Updated 3 years ago
- This repo is an exploratory experiment to enable frozen pretrained RWKV language models to accept speech modality input. We followed the …☆54Dec 23, 2024Updated last year
- We Speech Transcript based on LLM, in 300 lines of code.☆182Jun 20, 2025Updated last year
- Dynamic Mixing For Speech Processing (mix-on-the-fly)☆22Jul 19, 2022Updated 4 years ago
- Causal streaming adaptation of OpenAI Whisper for real-time transcription on small audio chunks.☆77Mar 31, 2026Updated 5 months ago