Deploy OmniVoice TTS model using TRT-LLM and Triton Inference Server on Modal.
☆18May 29, 2026Updated 3 months ago
Alternatives and similar repositories for omnivoice-trtllm
Users that are interested in omnivoice-trtllm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Triton kernel fusion & CUDA Graph optimization for OmniVoice inference — RMSNorm, SwiGLU, Norm+Residual, SageAttention☆63Jul 20, 2026Updated last month
- Browser-based text-to-speech powered by OmniVoice. Runs entirely locally via WebGPU and WebAssembly.☆17Jul 2, 2026Updated 2 months ago
- OpenAI-compatible HTTP server for OmniVoice text-to-speech☆83Aug 31, 2026Updated last week
- ☆18May 13, 2026Updated 4 months ago
- ☆26Jun 15, 2026Updated 2 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Open TTS implementation for ViiTorVoice-NAR☆20Jul 2, 2026Updated 2 months ago
- A simple and powerful tool for building Large Language Models from scratch【从零训练大模型】☆19Sep 29, 2025Updated 11 months ago
- Qwen3-TTS with nano vLLM-style optimizations for fast text-to-speech generation. Achieved 3x faster☆139Mar 3, 2026Updated 6 months ago
- ☆22May 2, 2026Updated 4 months ago
- an API server for indextts, allowing simple access without need to integrate actual code into your project☆18Oct 29, 2025Updated 10 months ago
- ☆40Jul 15, 2025Updated last year
- A local browser reading app powered by MOSS-TTS-Nano with in-browser ONNX inference☆60May 7, 2026Updated 4 months ago
- A tracing JIT compiler for PyTorch☆14Dec 11, 2021Updated 4 years ago
- How to use OpenAI API?☆13Nov 23, 2023Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- OpenMOSS pure C++ pipeline based on GGML☆72Aug 21, 2026Updated 3 weeks ago
- 2D Vector-Quantized Auto-Encoder for compression of Whole-Slide Images in Histopathology☆16Jul 18, 2024Updated 2 years ago
- [Eurographics 2025] ASMR: Adaptive Skeleton-Mesh Rigging and Skinning via 2D Generative Prior☆13Apr 22, 2025Updated last year
- ☆12Nov 7, 2024Updated last year
- [EMNLP 2024 Tutorial] Language Agents: Foundations, Prospects, and Risks☆10Nov 27, 2024Updated last year
- a Frontier Japanese Speech Generation net☆65May 15, 2025Updated last year
- A collection of tools for your LLMs that run on Modal☆25Feb 28, 2025Updated last year
- A python algorithm to change the pitch of the voice in real time☆13Dec 13, 2020Updated 5 years ago
- [ICCV 2025] STaR: Seamless Spatial-Temporal Aware Motion Retargeting with Penetration and Consistency Constraints☆19May 22, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Some LaTeX Tips for Writing Research Papers☆10May 30, 2016Updated 10 years ago
- Official implementation of "Weakly-supervised positional contrastive learning: application to cirrhosis classification", MICCAI 2023☆11Dec 16, 2025Updated 8 months ago
- LTX-2 is the first DiT-based audio-video foundation model that contains all core capabilities of modern video generation in one model: sy…☆31Updated this week
- a simple variational auto encoder with some exploration☆14Nov 22, 2024Updated last year
- Official implementation of the paper: “XSpecMesh: Quality-Preserving Auto-Regressive Mesh Generation Acceleration via Multi-Head Speculat…☆17Aug 7, 2025Updated last year
- ☆13Nov 13, 2025Updated 10 months ago
- Want to write papers but don't know where to start? Read on!☆11May 9, 2018Updated 8 years ago
- Proof-of-concept automated forensic dental identification system using Siamese Neural Networks for dental radiograph similarity analysis.…☆15Nov 19, 2025Updated 9 months ago
- ☆12Apr 3, 2026Updated 5 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Repository for evaluating Pegasus-1 and video-language foundation models☆14Nov 12, 2024Updated last year
- [MICCAI 2025] Hierarchical Self-Supervised Adversarial Training for Robust Vision Models in Histopathology☆12Jun 17, 2025Updated last year
- Code and models for the paper "An SpO2 Based Deep Learning Technique for Sleep Apnea Detection in Smart Watches."☆13Nov 18, 2024Updated last year
- ☆14Oct 3, 2018Updated 7 years ago
- Tutorial for Graph Neural Network at APBJC 2024.☆12Apr 21, 2025Updated last year
- Histopathology Feature Extractors (2024)☆14Jun 14, 2024Updated 2 years ago
- An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System☆148Nov 15, 2025Updated 9 months ago