stt websockect server using sherpa-onnx
☆59Feb 28, 2026Updated 6 months ago
Alternatives and similar repositories for stt-server
Users that are interested in stt-server are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- On-device speech AI runtime for ASR, TTS, VAD, and voice cloning. Python-simple, C++-native, GGUF-powered.☆27Aug 7, 2026Updated 3 weeks ago
- Port of Funasr's Paraformer model in C/C++☆43Jun 19, 2024Updated 2 years ago
- 一个基于 Sherpa-ONNX 的高性能语音识别服务,支持实时VAD(语音活动检测)、多语言语音识别和声纹识别功能。☆116Jan 4, 2026Updated 8 months ago
- SummerTTS 是一个基于C++的独立编译的中文和英文语音合成项目,可以本地运行不需要网络,而且没有额外的依赖,一键编译完成即可用于中文和英文的语音合成。SummerTTS is a standalone Chinese and English speech synt…☆25Aug 17, 2024Updated 2 years ago
- ☆25Mar 8, 2026Updated 5 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Streaming ASR and TTS based on FastAPI+ sherpa-onnx☆225Nov 2, 2025Updated 10 months ago
- Port of Funasr's Sense-voice model in C/C++☆573Dec 19, 2025Updated 8 months ago
- ☆50Jan 20, 2025Updated last year
- To convert CosyVoice model to ONNX☆16Dec 22, 2025Updated 8 months ago
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- A lightweight demo of FunASR-Nano using ONNX runtime.☆87Feb 25, 2026Updated 6 months ago
- low-latency realtime ASR based on FireRedASR☆61Jul 8, 2025Updated last year
- ☆13Mar 30, 2023Updated 3 years ago
- GTCRN(ncnn).☆19May 22, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- August智能体框架的相关功能包,已成功部署在Robonova机器人上☆20Aug 31, 2025Updated last year
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- A playground for experimenting with acoustic echo cancellation using a microphone, speaker, and ONNX.☆13Oct 22, 2024Updated last year
- Freeswitch Speech-To-Text module☆17Aug 25, 2026Updated last week
- A enterprise-grade Voice Activity Detector from modelscope and funasr.☆141Apr 26, 2023Updated 3 years ago
- The CPP version of Silero VAD: pre-trained enterprise-grade Voice Activity Detector☆23May 11, 2024Updated 2 years ago
- TEN VAD low-latency voice activity detection for real-time streaming, integrated with livekit-agents☆26Nov 13, 2025Updated 9 months ago
- real-time web visualizer for 3D gaussian splatting☆10Jan 31, 2025Updated last year
- Pseudo Streaming SenseVoice with Hotwords☆470Jun 15, 2026Updated 2 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- 一键将视频转换为优质小红书笔记,自动优化内容和配图;追加了可以读取本地视频的功能☆12Dec 22, 2024Updated last year
- A package used to test webrtc apm functions, such as aec, ns☆17Feb 21, 2019Updated 7 years ago
- A streaming audio reader, processor, and writer built on top of soundfile, and PyAV (bindings for FFmpeg)☆39Aug 27, 2026Updated last week
- Clean up noisy speech in real time with DPDFNet - open-source streaming speech enhancement for research, audio apps, and edge devices. In…☆139Jul 22, 2026Updated last month
- An open source chat bot architecture for voice/vision (and multimodal) assistants, local(CPU/GPU bound) and remote(I/O bound) to run.☆89Dec 28, 2025Updated 8 months ago
- Template for creating audio encoders compatible with X-ARES☆19Feb 11, 2026Updated 6 months ago
- We Speech Transcript based on LLM, in 300 lines of code.☆182Jun 20, 2025Updated last year
- Dynamic Mixing For Speech Processing (mix-on-the-fly)☆22Jul 19, 2022Updated 4 years ago
- Freeswitch ASR module☆22Jan 15, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Causal streaming adaptation of OpenAI Whisper for real-time transcription on small audio chunks.☆76Mar 31, 2026Updated 5 months ago
- ☆21Nov 28, 2025Updated 9 months ago
- X-ASR is a series of automatic speech recognition models based on the icefall framework, focusing on streaming ASR and low-latency deploy…☆174Jul 29, 2026Updated last month
- C++ version of openWakeWord☆44Jul 9, 2024Updated 2 years ago
- semantic tokenizer for speech and music☆20Jul 6, 2025Updated last year
- Daemonless background processes for Linux☆12Nov 28, 2024Updated last year
- This is a repository dedicated for pre-trained acoustic models of Hong Kong Cantonese and Cantonese forced alignment.☆29Nov 14, 2024Updated last year