3xMike / tritonserver-rsLinks
Rust crate for easy and efficient ML model inference
☆30Updated 6 months ago
Alternatives and similar repositories for tritonserver-rs
Users that are interested in tritonserver-rs are comparing it to the libraries listed below
Sorting:
- Rust library for running TensorRT accelerated deep learning models☆64Updated 4 years ago
- ☆40Updated last year
- Rust wrapper for Microsoft's ONNX Runtime (version 1.8)☆319Updated last year
- Rust client for the huggingface hub aiming for minimal subset of features over `huggingface-hub` python package☆254Updated last week
- An extension library to Candle that provides PyTorch functions not currently available in Candle☆40Updated last year
- A client library in Rust for Nvidia Triton.☆30Updated 2 years ago
- Asynchronous TensorRT for Rust.☆40Updated 4 months ago
- A rust port of pytorch dataloader☆30Updated last year
- A collection of optimisers for use with candle☆45Updated last month
- Low rank adaptation (LoRA) for Candle.☆169Updated 9 months ago
- Extract core logic from qdrant and make it available as a library.☆63Updated last year
- ☆33Updated last week
- Models and examples built with Burn☆331Updated last week
- implement llava using candle☆15Updated last year
- A framework for building high-performance real-time multiple object trackers☆256Updated 10 months ago
- Rust wrapper for Microsoft's ONNX Runtime with CUDA support (version 1.7)☆24Updated 3 years ago
- ONNX neural network inference engine☆282Updated 2 weeks ago
- Example of tch-rs on M1☆55Updated last year
- Rust bindings for OpenVINO™☆114Updated 3 weeks ago
- Inference Llama 2 in one file of pure Rust 🦀☆235Updated 2 years ago
- Rust language bindings for Faiss☆246Updated 2 months ago
- Savant Library with new generation primitives re-implemented in Rust☆19Updated this week
- GPU based FFT written in Rust and CubeCL☆29Updated last month
- Tutorial for Porting PyTorch Transformer Models to Candle (Rust)☆338Updated last year
- Rust implementation of the HNSW algorithm (Malkov-Yashunin)☆230Updated 2 months ago
- Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.☆578Updated last week
- ☆128Updated last year
- Playing around "Less Slow" coding practices in Rust, from numerical micro-kernels to coroutines, ranges, and polymorphic state machines☆122Updated 9 months ago
- Dataflow is a data processing library, primarily for machine learning.☆24Updated 2 years ago
- A Demo server serving Bert through ONNX with GPU written in Rust with <3☆42Updated 4 years ago