Utilizes ONNX Runtime for speech activity detection.
☆50Aug 17, 2026Updated last month
Alternatives and similar repositories for Voice-Activity-Detection-VAD-ONNX
Users that are interested in Voice-Activity-Detection-VAD-ONNX are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Uses ONNX Runtime for character role speaker identification.☆18Dec 28, 2025Updated 9 months ago
- Utilizes ONNX Runtime to transcribe audio into text.☆89Aug 26, 2026Updated last month
- Utilizes ONNX Runtime for audio denoising.☆138Aug 15, 2026Updated last month
- Export the STFT or ISTFT process in ONNX format.☆47Updated this week
- Utilizes ONNX Runtime for TTS model.☆70Sep 4, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SpeechDenoiser: Real-Time Speech Denoising with ONNX Welcome to SpeechDenoiser, a simple and effective solution for real-time speech den…☆119Aug 16, 2024Updated 2 years ago
- ☆24Jul 29, 2024Updated 2 years ago
- Export YOLO Series to ONNX format.☆73Aug 23, 2026Updated last month
- Running the F5-TTS by ONNX Runtime☆209Jul 29, 2026Updated 2 months ago
- Demonstrate Yolov9 model with Qualcomm Hexagon NPU and DirectML☆12Nov 27, 2024Updated last year
- 修复funasr中seaco-paraformer导出onnx后没有时间戳的bug☆25Sep 12, 2024Updated 2 years ago
- Running the F5-TTS by ONNX Runtime standalone with GUI☆29Dec 10, 2024Updated last year
- ☆53Mar 18, 2026Updated 6 months ago
- ☆19Jan 6, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- F5-TTS 推理加速,速度提升约4倍!☆123Jan 6, 2025Updated last year
- A library for adding punctuation into a text from ASR.☆20May 8, 2023Updated 3 years ago
- 音频处理小工具☆14Jun 4, 2026Updated 3 months ago
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆20Mar 12, 2026Updated 6 months ago
- Lightweight streaming Voice Activity Detection (VAD) tool with ONNX runtime☆24Mar 18, 2026Updated 6 months ago
- An example of a speech enhancement model deployed with TensorRT.☆90Mar 24, 2025Updated last year
- Demonstration of running a native LLM on Android device.☆265Aug 14, 2026Updated last month
- ☆17Feb 22, 2025Updated last year
- [Tiny KWS] SparkNet: Sparse Binarization for Fast Keyword Spotting☆20Aug 26, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This is the official implementation of PGUSE☆43Jun 7, 2025Updated last year
- YOLOv8安全帽工作服检测☆12Oct 13, 2023Updated 2 years ago
- ☆19Mar 15, 2024Updated 2 years ago
- 目标检测算法主要包括:两类two-stage和one-stage 一类是two-stage,two-stage检测算法将检测问题划分为两个阶段,首先产生候选区域(region proposals),然后对候选区域分类(一般还需要对位置精修),这一类的典型代表是R-CNN…☆16Sep 5, 2021Updated 5 years ago
- 用 onnx 和 gguf 格式混合运行 Fun-ASR-Nano 模型全流程☆161May 5, 2026Updated 4 months ago
- 达摩fsmn vad c++推理服务☆17Apr 17, 2023Updated 3 years ago
- The official implementation of GTCRN, an ultra-lightweight SE model.☆748Aug 3, 2026Updated 2 months ago
- Voice activity engine benchmark framework☆23Jan 14, 2026Updated 8 months ago
- 基于飞桨pphuman中跨镜头跟踪的改进,主要是实现两路推流实时跟踪☆16Aug 14, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A SOTA Industrial-Grade Voice Activity Detection & Audio Event Detection, supporting 100+ languages, outperforming Silero-VAD, TEN-VAD, F…☆542May 6, 2026Updated 4 months ago
- VMX Codec☆24Sep 23, 2026Updated last week
- 基于ultralytics训练的行人跌倒检测模型☆20Jul 10, 2023Updated 3 years ago
- Context-Sensitive Neural Spelling Checker☆20Sep 25, 2024Updated 2 years ago
- Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR be…☆1,998Feb 25, 2026Updated 7 months ago
- Unofficial implementation of 'Large-Scale Text-to-Image Model with Inpainting is a Zero-Shot Subject-Driven Image Generator'☆10Dec 10, 2024Updated last year
- BUD-E (Buddy) is an open-source voice assistant framework that facilitates seamless interaction with AI models and APIs, enabling the cre…☆23Oct 10, 2024Updated last year