OpenAI-compatible HTTP server for OmniVoice text-to-speech
☆82Aug 31, 2026Updated last week
Alternatives and similar repositories for omnivoice-server
Users that are interested in omnivoice-server are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Browser-based text-to-speech powered by OmniVoice. Runs entirely locally via WebGPU and WebAssembly.☆16Jul 2, 2026Updated 2 months ago
- ☆22May 27, 2026Updated 3 months ago
- Deploy OmniVoice TTS model using TRT-LLM and Triton Inference Server on Modal.☆18May 29, 2026Updated 3 months ago
- TTS-Story is a web-based multi‑voice TTS studio for turning tagged scripts into audiobooks—featuring full speaker management, chunk revie…☆332Aug 31, 2026Updated last week
- Evaluation tool used in the BigVSAN paper☆14Mar 22, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆14Oct 27, 2025Updated 10 months ago
- Local AI text-to-speech with voice cloning and voice design, powered by GGML. C++17 port of OmniVoice (k2-fsa/OmniVoice). 646 languages, …☆167Updated this week
- High-Quality Voice Cloning TTS for 600+ Languages☆10,351Aug 31, 2026Updated last week
- Triton kernel fusion & CUDA Graph optimization for OmniVoice inference — RMSNorm, SwiGLU, Norm+Residual, SageAttention☆62Jul 20, 2026Updated last month
- magicspeech competition recipe☆18Jun 29, 2020Updated 6 years ago
- A beginner-friendly inference to finetune & run inference on open TTS models 🗣️☆30Feb 4, 2026Updated 7 months ago
- Echo-TTS inference codebase☆224Dec 5, 2025Updated 9 months ago
- Echo-TTS OpenAI Compatible Speech Endpoint w/ Streaming☆29Apr 5, 2026Updated 5 months ago
- Application to convert Vietnamese sign language via camera to audio☆18Aug 26, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Generative Motion Latent Flow Matching for Audio-driven Talking Portrait☆33Sep 10, 2025Updated 11 months ago
- Lightweight streaming Voice Activity Detection (VAD) tool with ONNX runtime☆24Mar 18, 2026Updated 5 months ago
- ☆26Jun 15, 2026Updated 2 months ago
- Context Query language for Agents☆65Jul 31, 2026Updated last month
- Real-time text-to-speech with Qwen3-TTS☆1,339Aug 25, 2026Updated 2 weeks ago
- semantic tokenizer for speech and music☆20Jul 6, 2025Updated last year
- A streaming audio reader, processor, and writer built on top of soundfile, and PyAV (bindings for FFmpeg)☆39Aug 27, 2026Updated last week
- ☆44Nov 19, 2025Updated 9 months ago
- Speech Dereverberation using weighted prediction error☆11Dec 22, 2019Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Scaled diffusion transformer for text-to-speech synthesis (DiT + T5Gemma2 conditioning, TorchTitan & Megatron backends, tested up to 1024…☆24Mar 29, 2026Updated 5 months ago
- Local text-to-speech in your browser with Piper TTS☆26Aug 13, 2025Updated last year
- Signal generator designed with Nexy4 FPGA☆13May 14, 2023Updated 3 years ago
- ☆15Jun 6, 2025Updated last year
- 使用肤色颜色空间建模+连通域处理及分析和Harr-cascade 方法进行人脸检测。1建立多种肤色模型,结合数学形态学滤波,完成人脸检测; 2利用Matlab 自带的计算机视觉系统工具箱实现单人及多人的人脸检测。☆13Nov 23, 2018Updated 7 years ago
- For audio visualization and playback in Jupyter notebooks.☆19Aug 27, 2026Updated last week
- Yet Another (LLM) Web UI, made with Gemini☆12Dec 25, 2024Updated last year
- speech-dereverberation-using-GANs☆13Jan 28, 2019Updated 7 years ago
- Code for vec2wav 2.0, a speech token vocoder for VC. Paper: https://arxiv.org/abs/2409.01995☆79Dec 3, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This challenge focuses on evaluating speech recognition and semantic understanding capabilities of AI glasses in complex real-world envir…☆18Aug 22, 2026Updated 2 weeks ago
- High-level API for tar-based dataset☆12Feb 3, 2024Updated 2 years ago
- A car Heads Up Display built using a RGB LED strip and a Teensy microcontroller☆10Jul 5, 2017Updated 9 years ago
- Hand Writing Digital Recognization Based on FPGA, we desiged a SoC embeded a Cortex M3 core and other peripherals,this SoC run a CNN. The…☆14Mar 30, 2023Updated 3 years ago
- Nexys 4 DDR Artix-7☆11Jun 15, 2018Updated 8 years ago
- ☆39Jun 9, 2025Updated last year
- A family of image super-resolution models with purrfect pixels.☆16Apr 29, 2026Updated 4 months ago