Triton kernel fusion & CUDA Graph optimization for OmniVoice inference — RMSNorm, SwiGLU, Norm+Residual, SageAttention
☆63Jul 20, 2026Updated 2 months ago
Alternatives and similar repositories for omnivoice-triton
Users that are interested in omnivoice-triton are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Deploy OmniVoice TTS model using TRT-LLM and Triton Inference Server on Modal.☆18May 29, 2026Updated 4 months ago
- ☆20May 13, 2026Updated 4 months ago
- ☆26Mar 1, 2026Updated 7 months ago
- Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation☆76Jun 16, 2026Updated 3 months ago
- ☆26Jun 15, 2026Updated 3 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Triton kernel fusion for Qwen3-TTS 1.7B inference acceleration — RMSNorm, SwiGLU, M-RoPE, Norm+Residual☆107Jun 29, 2026Updated 3 months ago
- ☆315Sep 24, 2026Updated last week
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆18Jun 27, 2026Updated 3 months ago
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆35Sep 25, 2025Updated last year
- A Python library for text normalization, specifically designed for Vietnamese and English text processing. This library provides comprehe…☆14Mar 30, 2025Updated last year
- Triton kernels for Flux☆23Jul 7, 2025Updated last year
- Graph model execution API for Candle☆18Jul 27, 2025Updated last year
- OpenAI-compatible HTTP server for OmniVoice text-to-speech☆89Aug 31, 2026Updated last month
- Official PyTorch inference code for the Interspeech 2025 paper: Efficient Speech Enhancement via Embeddings from Pre-trained Generative A…☆82Jun 16, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A high-performance batch audio transcription tool using nvidia/parakeet-tdt-0.6b-v2 to generate accurate, well-segmented SRT subtitles, w…☆18Dec 9, 2025Updated 9 months ago
- An end-to-end ASR model, transcribing spoken Chinese to formal text.☆23Jun 26, 2026Updated 3 months ago
- ☆37Jan 6, 2026Updated 8 months ago
- Eureka-Audio: A 1.7B lightweight audio–language model that matches 7B–30B models on ASR, audio understanding, and paralinguistic reasonin…☆46Apr 11, 2026Updated 5 months ago
- Browser-based text-to-speech powered by OmniVoice. Runs entirely locally via WebGPU and WebAssembly.☆17Jul 2, 2026Updated 3 months ago
- A demo-level low-latency, high-throughput inference engine for whisper☆20Nov 9, 2025Updated 10 months ago
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated last year
- fd-sds☆21Apr 8, 2026Updated 5 months ago
- Python runtime for WeTextProcessing (does not depend on Pynini)☆61Sep 9, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- SoTA open-source TTS☆26Jul 8, 2025Updated last year
- For audio visualization and playback in Jupyter notebooks.☆19Aug 27, 2026Updated last month
- A Comprehensive Speech Processing Algorithms Library for research and production use☆18Oct 25, 2025Updated 11 months ago
- IndexTTS Fine-tuning notebooks☆140Jun 17, 2025Updated last year
- PyTorch -> ONNX☆17Oct 18, 2025Updated 11 months ago
- High-Quality Voice Cloning TTS for 600+ Languages☆14,125Updated this week
- Unofficial fairseq-free PyTorch implementation of UTMOS (v1, 2022), matching the original system.☆36Jun 6, 2026Updated 3 months ago
- The first open source triton inference engine for Stable Diffusion, specifically for sdxl☆12Nov 27, 2023Updated 2 years ago
- X-Voice☆184Sep 24, 2026Updated last week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Streaming Text to Speech Web UI☆22May 6, 2024Updated 2 years ago
- Rust standalone inference of Namo-500M series models. Extremly tiny, runing VLM on CPU.☆24Mar 12, 2025Updated last year
- ☆40Jul 21, 2026Updated 2 months ago
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching☆1,076Dec 2, 2025Updated 10 months ago
- On-device VAD / streaming STT / TTS / diarization in C++17 (ONNX + LiteRT) with a voice-agent pipeline. Linux, Windows, Android.☆90Sep 15, 2026Updated 2 weeks ago
- Official repository for Mamba-based Segmentation Model for Speaker Diarization☆47May 13, 2025Updated last year
- Pure-PyTorch Parakeet TDT inference☆54Mar 10, 2026Updated 6 months ago