Triton kernel fusion & CUDA Graph optimization for OmniVoice inference — RMSNorm, SwiGLU, Norm+Residual, SageAttention
☆59Jul 20, 2026Updated 2 weeks ago
Alternatives and similar repositories for omnivoice-triton
Users that are interested in omnivoice-triton are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Deploy OmniVoice TTS model using TRT-LLM and Triton Inference Server on Modal.☆17May 29, 2026Updated 2 months ago
- ☆18May 13, 2026Updated 2 months ago
- ☆22Mar 1, 2026Updated 5 months ago
- Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation☆68Jun 16, 2026Updated last month
- ☆25Jun 15, 2026Updated last month
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Triton kernel fusion for Qwen3-TTS 1.7B inference acceleration — RMSNorm, SwiGLU, M-RoPE, Norm+Residual☆101Jun 29, 2026Updated last month
- ☆282Updated this week
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆15Jun 27, 2026Updated last month
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆35Sep 25, 2025Updated 10 months ago
- A Python library for text normalization, specifically designed for Vietnamese and English text processing. This library provides comprehe…☆14Mar 30, 2025Updated last year
- Triton kernels for Flux☆23Jul 7, 2025Updated last year
- Graph model execution API for Candle☆18Jul 27, 2025Updated last year
- An end-to-end ASR model, transcribing spoken Chinese to formal text.☆21Jun 26, 2026Updated last month
- OpenAI-compatible HTTP server for OmniVoice text-to-speech☆75Jun 26, 2026Updated last month
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Official PyTorch inference code for the Interspeech 2025 paper: Efficient Speech Enhancement via Embeddings from Pre-trained Generative A…☆81Jun 16, 2025Updated last year
- A high-performance batch audio transcription tool using nvidia/parakeet-tdt-0.6b-v2 to generate accurate, well-segmented SRT subtitles, w…☆18Dec 9, 2025Updated 7 months ago
- ☆37Jan 6, 2026Updated 6 months ago
- Eureka-Audio: A 1.7B lightweight audio–language model that matches 7B–30B models on ASR, audio understanding, and paralinguistic reasonin…☆40Apr 11, 2026Updated 3 months ago
- Browser-based text-to-speech powered by OmniVoice. Runs entirely locally via WebGPU and WebAssembly.☆16Jul 2, 2026Updated last month
- A demo-level low-latency, high-throughput inference engine for whisper☆20Nov 9, 2025Updated 8 months ago
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated 11 months ago
- Implementation of the OmniVoice inference model from k2-fsa on Rust☆26Updated this week
- fd-sds☆21Apr 8, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Python runtime for WeTextProcessing (does not depend on Pynini)☆53Updated this week
- SoTA open-source TTS☆26Jul 8, 2025Updated last year
- For audio visualization and playback in Jupyter notebooks.☆18Nov 25, 2025Updated 8 months ago
- High-Quality Voice Cloning TTS for 600+ Languages☆8,742Updated this week
- A Comprehensive Speech Processing Algorithms Library for research and production use☆18Oct 25, 2025Updated 9 months ago
- IndexTTS Fine-tuning notebooks☆139Jun 17, 2025Updated last year
- PyTorch -> ONNX☆17Oct 18, 2025Updated 9 months ago
- Unofficial fairseq-free PyTorch implementation of UTMOS (v1, 2022), matching the original system.☆35Jun 6, 2026Updated last month
- ✅ CodeBlock's frontend development course starter file ❤️☆14Nov 13, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- The first open source triton inference engine for Stable Diffusion, specifically for sdxl☆12Nov 27, 2023Updated 2 years ago
- X-Voice☆177Jun 5, 2026Updated last month
- Streaming Text to Speech Web UI☆22May 6, 2024Updated 2 years ago
- Rust standalone inference of Namo-500M series models. Extremly tiny, runing VLM on CPU.☆24Mar 12, 2025Updated last year
- On-device VAD / streaming STT / TTS / diarization in C++17 (ONNX + LiteRT) with a voice-agent pipeline. Linux, Windows, Android.☆61Updated this week
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching☆1,029Dec 2, 2025Updated 8 months ago
- ☆30Jul 21, 2026Updated 2 weeks ago