octoml / triton-client-rs
A client library in Rust for Nvidia Triton.
☆24Updated last year
Related projects ⓘ
Alternatives and complementary repositories for triton-client-rs
- Asynchronous CUDA for Rust.☆27Updated 2 weeks ago
- Modular Rust transformer/LLM library using Candle☆36Updated 6 months ago
- ☆19Updated 4 months ago
- ☆25Updated last year
- Rust library for whisper.cpp compatible Mel spectrograms☆58Updated 2 months ago
- Implementing the BitNet model in Rust☆28Updated 7 months ago
- Rust implementation of Huggingface transformers pipelines using onnxruntime backend with bindings to C# and C.☆34Updated last year
- ESRGAN implemented in rust with candle☆14Updated 11 months ago
- A Demo server serving Bert through ONNX with GPU written in Rust with <3☆39Updated 3 years ago
- ☆17Updated last month
- Sample Python extension using Rust/PyO3/tch to interact with PyTorch☆32Updated 9 months ago
- A Fish Speech implementation in Rust, with Candle.rs☆45Updated this week
- Your one stop CLI for ONNX model analysis.☆45Updated 2 years ago
- LLaMa 7b with CUDA acceleration implemented in rust. Minimal GPU memory needed!☆101Updated last year
- ☆22Updated this week
- Experimental ONNX implementation for WASI NN.☆47Updated 3 years ago
- A diffusers API in Burn (Rust)☆15Updated 4 months ago
- python bindings for symphonia/opus - read various audio formats from python and write opus files☆21Updated 2 months ago
- Example of tch-rs on M1☆51Updated 8 months ago
- Dataflow is a data processing library, primarily for machine learning.☆19Updated last year
- Low rank adaptation (LoRA) for Candle.☆127Updated 3 months ago
- Rust client for the huggingface hub aiming for minimal subset of features over `huggingface-hub` python package☆153Updated 2 months ago
- Tantivy directory implementation backed by object_store☆27Updated 9 months ago
- Simple dependency injection framework for Python☆20Updated 6 months ago
- 8-bit floating point types for Rust☆39Updated last month
- Fast serverless LLM inference, in Rust.☆22Updated 2 weeks ago
- ☆12Updated 10 months ago
- A collection of optimisers for use with candle☆31Updated this week
- A high-performance constrained decoding engine based on context free grammar in Rust☆40Updated last week
- implement llava using candle☆13Updated 5 months ago