Real-time speech recognition and voice activity detection (VAD) using next-gen Kaldi with ncnn without Internet connection. Support iOS, Android, Linux, macOS, Windows, Raspberry Pi, VisionFive2, LicheePi4A etc.
☆1,784Oct 20, 2025Updated 11 months ago
Alternatives and similar repositories for sherpa-ncnn
Users that are interested in sherpa-ncnn are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Speech-to-text server framework with next-gen Kaldi☆995Updated this week
- Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime…☆14,907Updated this week
- ☆1,504Jul 16, 2026Updated 2 months ago
- FSA/FST algorithms, differentiable, with PyTorch compatibility.☆1,361Jul 11, 2026Updated 2 months ago
- vits Android部署☆353Mar 31, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 这是一个用C++实现ASR推理的项目,它依赖很少,安装也很简单,推理速度很快,在树莓派4B等ARM平台也可以流畅的运行。 支持的模型是由Google的Transformer模型中优化而来,数据集是开源wenetspeech(10000+小时)或阿里私有数据集(60000+小…☆555Mar 19, 2023Updated 3 years ago
- Kaldi-compatible online fbank extractor without external dependencies☆155Oct 9, 2025Updated 11 months ago
- libvits-ncnn is an ncnn implementation of the VITS library that enables cross-platform GPU-accelerated speech synthesis.🎙️💻☆62May 6, 2023Updated 3 years ago
- Python wrapper for OpenFST and its extensions from Kaldi. Also support reading/writing ark/scp files☆56Apr 9, 2026Updated 5 months ago
- ncnn HiFi-GAN☆30Sep 29, 2024Updated last year
- Colab notebooks for Next-gen Kaldi☆31Oct 12, 2025Updated 11 months ago
- Production First and Production Ready End-to-End Speech Recognition Toolkit☆5,239Sep 7, 2026Updated 2 weeks ago
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆20,474Updated this week
- sherpa with mlx☆15Aug 2, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 📣 商用 级开源语音自动识别程序库,开箱即用,全平台支持,中英文混合识别。A Cross-platform implementation of ASR inference. It's based on ONNXRuntime and FunASR. We provide …☆610May 15, 2024Updated 2 years ago
- Android paddleocr demo infer by ncnn☆210Jul 23, 2024Updated 2 years ago
- ncnn is a high-performance neural network inference framework optimized for the mobile platform☆23,861Updated this week
- Production First and Production Ready End-to-End Keyword Spotting Toolkit☆764Jul 23, 2026Updated 2 months ago
- GPT2⚡NCNN⚡中文对话⚡x86⚡Android☆82Mar 25, 2022Updated 4 years ago
- Tools for handling multimodal data in machine learning projects.☆1,152Sep 11, 2026Updated last week
- Moved to https://github.com/k2-fsa/icefall☆146Oct 13, 2022Updated 3 years ago
- Kaldi-compatible online & offline feature extraction with PyTorch, supporting CUDA, batch processing, chunk processing, and autograd - P…☆220Jul 10, 2026Updated 2 months ago
- Decoders from Kaldi using OpenFst☆35Apr 10, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The minimal opencv for Android, iOS, ARM Linux, Windows, Linux, MacOS, HarmonyOS, WebAssembly, watchOS, tvOS, visionOS☆3,361Jul 12, 2026Updated 2 months ago
- ☆82Jun 29, 2023Updated 3 years ago
- DragGan in NCNN with c++☆52Oct 5, 2023Updated 2 years ago
- Silero VAD: pre-trained enterprise-grade Voice Activity Detector☆10,281Updated this week
- some ncnn demos of FunASR☆29Sep 23, 2024Updated 2 years ago
- awesome AI models with NCNN, and how they were converted ✨✨✨☆282Jul 7, 2023Updated 3 years ago
- PyTorch Neural Network eXchange☆714Updated this week
- ncnn version of CodeFormer☆109Mar 9, 2023Updated 3 years ago
- Production First and Production Ready End-to-End Text-to-Speech Toolkit☆417Nov 20, 2025Updated 10 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A Deeplearn Model to rec table in photo with ncnn. 一个深度学习模型用于检测图片中的表格 画像内のテーブルを検出するためのディープラーニング モデル☆20Mar 2, 2025Updated last year
- A converter for llama2.c legacy models to ncnn models.☆79Dec 17, 2023Updated 2 years ago
- Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio…☆9,349Updated this week
- Inference TinyLlama models on ncnn☆24Aug 15, 2023Updated 3 years ago
- ncnn demo of (文档矫正)DocTr: Document Image Transformer for Geometric Unwarping and Illumination Correction☆46Jan 9, 2022Updated 4 years ago
- Hybrid Flow Matching and GAN with Multi-Resolution Network for Few-Step High-Fidelity Audio Generation☆149Mar 8, 2026Updated 6 months ago
- a naive NCNN implementation of GFPGAN aims at developing Practical Algorithms for Real-world Face Restoration☆85Dec 29, 2021Updated 4 years ago