☆20May 13, 2026Updated 4 months ago
Alternatives and similar repositories for Omnivoice-onnx
Users that are interested in Omnivoice-onnx are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A family of image super-resolution models with purrfect pixels.☆16Apr 29, 2026Updated 5 months ago
- Code for CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment, Interspeech 2026.☆22Aug 14, 2026Updated last month
- ☆53Mar 18, 2026Updated 6 months ago
- Local AI text-to-speech with voice cloning and voice design, powered by GGML. C++17 port of OmniVoice (k2-fsa/OmniVoice). 646 languages, …☆182Updated this week
- Utilizes ONNX Runtime for TTS model.☆70Sep 4, 2026Updated 3 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for training and evaluation on the "Industrial Language-Image Dataset (ILID)".☆10Jun 4, 2025Updated last year
- Keyword Spotting using BCResNet and Arcface Loss☆13Jan 28, 2022Updated 4 years ago
- SenseVoice-Small 导出为 ONNX,支持热词注入,在 CTC 的输空间中通过路径匹配,1ms 内实现热词替换☆32Jun 3, 2026Updated 4 months ago
- 最极速的Qwen3-TTS推理方案。将 Qwen3-TTS 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆190Sep 19, 2026Updated 2 weeks ago
- A small and simple example showing how to run Qwen3-ASR with ONNX Runtime.☆40Apr 8, 2026Updated 5 months ago
- A GUI proxy client in GTK+3. Based on Clash.☆14Feb 12, 2019Updated 7 years ago
- The dataset construction pipeline for WordVoice-5A☆19Jul 17, 2026Updated 2 months ago
- Python binding of bark.cpp via Ctypes☆10Jan 1, 2025Updated last year
- SynTTS-Commands is a large-scale, multilingual (English & Chinese) synthetic speech command dataset designed for low-power Keyword Spotti…☆18Feb 5, 2026Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- End to End Speech to Speech with Emotion System☆15Feb 6, 2025Updated last year
- Python runtime for WeTextProcessing (does not depend on Pynini)☆61Sep 9, 2026Updated 3 weeks ago
- This node is base on VisualCloze method, A Universal Image Generation Framework via Visual In-Context Learning☆11May 21, 2025Updated last year
- ComfyUI-KokoroTTS: A text-to-speech model that utilizes the Kokoro TTS framework to convert text into natural-sounding speech. It suppor…☆13Apr 18, 2025Updated last year
- A TensorFlow port of Pyramid Stereo Matching Network (CVPR2018)☆10Dec 8, 2018Updated 7 years ago
- A universal phone recognizer that can transcribe speech in 100+ languages into IPA☆46Updated this week
- Deploy OmniVoice TTS model using TRT-LLM and Triton Inference Server on Modal.☆18May 29, 2026Updated 4 months ago
- This repo contains both matlab and python implementation of using TPS to perform warp given a set of correspondence points.☆20Aug 25, 2021Updated 5 years ago
- Recognize speech to text, and convert text matching speech to srt format subtitle file☆13Nov 20, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Offline audiobook reader for EPUB, PDF, and TXT with local OmniVoice TTS, character voices, and synced highlighting.☆23Jul 21, 2026Updated 2 months ago
- ☆26Jun 15, 2026Updated 3 months ago
- Open TTS implementation for ViiTorVoice-NAR☆21Jul 2, 2026Updated 3 months ago
- A simple and powerful tool for building Large Language Models from scratch【从零训练大模型】☆19Sep 29, 2025Updated last year
- ☆22May 2, 2026Updated 5 months ago
- 前端抠图-canvas实现☆13May 4, 2020Updated 6 years ago
- A demo of 2D WebView for WebGL.☆16Apr 22, 2026Updated 5 months ago
- A Python library for text normalization, specifically designed for Vietnamese and English text processing. This library provides comprehe…☆14Mar 30, 2025Updated last year
- IndexTTS Fine-tuning notebooks☆140Jun 17, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 小红书的flux版本的透明图生成(layerdiffuse),支持文生图和图生图☆18Mar 17, 2025Updated last year
- PersonaTalk Hack☆15Jan 10, 2025Updated last year
- 🚀 A centralized launcher for managing and running BAT and PS1 scripts.☆18Jun 20, 2026Updated 3 months ago
- SuperDex brings together a purpose-built physics engine, robotics authoring tools, and a scalable reinforcement learning interface in a u…☆710Updated this week
- https://github.com/xie-lab-ml/Golden-Noise-for-Diffusion-Models for ComfyUI☆18Dec 10, 2024Updated last year
- 本项目包含一个 Python 脚本,用于分离双人(或多人)对话播客音频文件中的不同说话人语音。它利用 `pyannote.audio` 库进行说话人日志分析(Speaker Diarization),找出“谁在什么时候说话”,并将每个说话人的语音片段提取到单独的音轨中。☆18Apr 30, 2025Updated last year
- Node.js implementation binding for the RWKV.cpp module☆22Aug 2, 2023Updated 3 years ago