On-device speech AI runtime for ASR, TTS, VAD, and voice cloning. Python-simple, C++-native, GGUF-powered.
☆21Jul 15, 2026Updated this week
Alternatives and similar repositories for RapidSpeech.cpp
Users that are interested in RapidSpeech.cpp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- stt websockect server using sherpa-onnx☆56Feb 28, 2026Updated 4 months ago
- mnn asr demo.☆27Mar 24, 2025Updated last year
- silero-vad pytorch implement☆38Nov 23, 2024Updated last year
- ☆25Mar 8, 2026Updated 4 months ago
- Lightweight streaming Voice Activity Detection (VAD) tool with ONNX runtime☆22Mar 18, 2026Updated 4 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Converting Chinese sentences into pinyin sequences, implemented in C++, very fast and easy to deploy.☆23Jan 5, 2026Updated 6 months ago
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- Streaming Vocos☆31Jun 10, 2025Updated last year
- High-performance C++ voice interaction framework powered by ONNXRuntime and LLaMA.cpp. Features AEC, VAD, ASR, TTS, LLM, and MCP integrat…☆51Mar 5, 2026Updated 4 months ago
- noise reduction☆17Jul 3, 2024Updated 2 years ago
- Port of Funasr's Paraformer model in C/C++☆43Jun 19, 2024Updated 2 years ago
- A package used to test webrtc apm functions, such as aec, ns☆17Feb 21, 2019Updated 7 years ago
- Uses the excellent silero VAD with onnxruntime C api for fast detection of audio segments with speech☆16Sep 20, 2024Updated last year
- Port of Funasr's Sense-voice model in C/C++☆568Dec 19, 2025Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- CTC decoder with hotwords for ASR.☆37Jun 15, 2026Updated last month
- This is a depth-anything-v2 onnxruntime inference by cpp☆16Sep 2, 2024Updated last year
- ☆14Apr 4, 2025Updated last year
- This repository features a from-scratch implementation of a neural network using CUDA and C. The primary goal of this project is to lever…☆12Mar 20, 2025Updated last year
- Unofficial implementation of wavenext vocoder☆59Aug 28, 2024Updated last year
- A unified tokenizer that is capable of both extracting semantic information and enabling high-fidelity audio reconstruction.☆145Sep 19, 2025Updated 10 months ago
- FlowMirror-HydraVox — A natively accelerated multi-head autoregressive TTS system derived from CosyVoice 3.0. It predicts multiple tokens…☆49Feb 17, 2026Updated 5 months ago
- Fun-ASR is an end-to-end speech recognition large model launched by Tongyi Lab.☆107Jul 7, 2026Updated last week
- 轻量级文本转语音工具,面向本地快速推理。A text-to-speech framework for fast and high-quality speech synthesis.☆55May 29, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ggml implementation of embedding models including SentenceTransformer and BGE☆66Dec 21, 2023Updated 2 years ago
- source code of EfficientTTS 2☆21Feb 18, 2024Updated 2 years ago
- A lightweight demo of FunASR-Nano using ONNX runtime.☆83Feb 25, 2026Updated 4 months ago
- ☆24Jul 29, 2024Updated last year
- SummerTTS 是一个基于C++的独立编译的中文和英文语音合成项目,可以本地运行不需要网络,而且没有额外的依赖,一键编译完成即可用于中文和英文的语音合成。SummerTTS is a standalone Chinese and English speech synt…☆25Aug 17, 2024Updated last year
- An end-to-end ASR model, transcribing spoken Chinese to formal text.☆20Jun 26, 2026Updated 3 weeks ago
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- The official code of OCT2Former for Retinal OCT-Angiography vessel segmentation☆17May 8, 2023Updated 3 years ago
- SenseVoice-python: A enterprise-grade open source multi-language asr system from funasr opensource with onnxruntime☆114Jun 12, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- tomato, a plant from the solanaceae family - tox built-in☆12Jul 9, 2026Updated last week
- A python library for voice activity detection (VAD) for speech/non-speech segmentation.☆88Sep 7, 2022Updated 3 years ago
- 完全独立编译 AEC, AGC, NS, VAD in WebRTC☆22Jul 8, 2019Updated 7 years ago
- Reimplementation of Miipher☆30Aug 16, 2023Updated 2 years ago
- DiffSinger dataset processing tools, including audio processing, labeling.☆85Jun 8, 2026Updated last month
- We Speech Transcript based on LLM, in 300 lines of code.☆182Jun 20, 2025Updated last year
- Distillation of Self-Supervised Representation-Based Speech Quality Assessment☆49May 15, 2025Updated last year