Triton kernel fusion & CUDA Graph optimization for OmniVoice inference — RMSNorm, SwiGLU, Norm+Residual, SageAttention
☆63Jul 20, 2026Updated last month
Alternatives and similar repositories for omnivoice-triton
Users that are interested in omnivoice-triton are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Deploy OmniVoice TTS model using TRT-LLM and Triton Inference Server on Modal.☆18May 29, 2026Updated 3 months ago
- ☆18May 13, 2026Updated 4 months ago
- ☆24Mar 1, 2026Updated 6 months ago
- Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation☆75Jun 16, 2026Updated 2 months ago
- ☆26Jun 15, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Triton kernel fusion for Qwen3-TTS 1.7B inference acceleration — RMSNorm, SwiGLU, M-RoPE, Norm+Residual☆105Jun 29, 2026Updated 2 months ago
- ☆303Sep 2, 2026Updated last week
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆17Jun 27, 2026Updated 2 months ago
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆35Sep 25, 2025Updated 11 months ago
- A Python library for text normalization, specifically designed for Vietnamese and English text processing. This library provides comprehe…☆14Mar 30, 2025Updated last year
- Graph model execution API for Candle☆18Jul 27, 2025Updated last year
- An end-to-end ASR model, transcribing spoken Chinese to formal text.☆22Jun 26, 2026Updated 2 months ago
- OpenAI-compatible HTTP server for OmniVoice text-to-speech☆83Aug 31, 2026Updated last week
- Official PyTorch inference code for the Interspeech 2025 paper: Efficient Speech Enhancement via Embeddings from Pre-trained Generative A…☆82Jun 16, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A high-performance batch audio transcription tool using nvidia/parakeet-tdt-0.6b-v2 to generate accurate, well-segmented SRT subtitles, w…☆18Dec 9, 2025Updated 9 months ago
- ☆37Jan 6, 2026Updated 8 months ago
- Eureka-Audio: A 1.7B lightweight audio–language model that matches 7B–30B models on ASR, audio understanding, and paralinguistic reasonin…☆45Apr 11, 2026Updated 5 months ago
- Browser-based text-to-speech powered by OmniVoice. Runs entirely locally via WebGPU and WebAssembly.☆17Jul 2, 2026Updated 2 months ago
- A demo-level low-latency, high-throughput inference engine for whisper☆20Nov 9, 2025Updated 10 months ago
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated last year
- Implementation of the OmniVoice inference model from k2-fsa on Rust☆29Updated this week
- fd-sds☆21Apr 8, 2026Updated 5 months ago
- Python runtime for WeTextProcessing (does not depend on Pynini)☆61Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- SoTA open-source TTS☆26Jul 8, 2025Updated last year
- For audio visualization and playback in Jupyter notebooks.☆19Aug 27, 2026Updated 2 weeks ago
- This tool allows local LLM usage that can automate tasks without human interventention. The agent can call itself recursively and work on…☆20May 5, 2025Updated last year
- IndexTTS Fine-tuning notebooks☆139Jun 17, 2025Updated last year
- PyTorch -> ONNX☆17Oct 18, 2025Updated 10 months ago
- High-Quality Voice Cloning TTS for 600+ Languages☆12,689Updated this week
- Unofficial fairseq-free PyTorch implementation of UTMOS (v1, 2022), matching the original system.☆35Jun 6, 2026Updated 3 months ago
- X-Voice☆182Aug 5, 2026Updated last month
- Streaming Text to Speech Web UI☆22May 6, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Rust standalone inference of Namo-500M series models. Extremly tiny, runing VLM on CPU.☆24Mar 12, 2025Updated last year
- On-device VAD / streaming STT / TTS / diarization in C++17 (ONNX + LiteRT) with a voice-agent pipeline. Linux, Windows, Android.☆83Sep 2, 2026Updated last week
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching☆1,061Dec 2, 2025Updated 9 months ago
- ☆40Jul 21, 2026Updated last month
- web application, powered by Python Flask and OpenAI GPT-3, designed to generate exceptional AI-generated content for a wide range of appl…☆13Feb 7, 2023Updated 3 years ago
- Official repository for Mamba-based Segmentation Model for Speaker Diarization☆47May 13, 2025Updated last year
- Pure-PyTorch Parakeet TDT inference☆54Mar 10, 2026Updated 6 months ago