huggingface / candle-paged-attentionLinks
☆12Updated 2 years ago
Alternatives and similar repositories for candle-paged-attention
Users that are interested in candle-paged-attention are comparing it to the libraries listed below
Sorting:
- ☆19Updated 2 weeks ago
- A collection of optimisers for use with candle☆45Updated 3 weeks ago
- implement llava using candle☆15Updated last year
- Rust crate for some audio utilities☆26Updated 10 months ago
- ☆18Updated last year
- CLI utility to inspect and explore .safetensors and .gguf files☆39Updated 2 months ago
- Read and write tensorboard data using Rust☆24Updated last year
- ☆28Updated 2 years ago
- Graph model execution API for Candle☆17Updated 5 months ago
- Your one stop CLI for ONNX model analysis.☆47Updated 3 years ago
- ☆13Updated 3 weeks ago
- Simple (fast) transformer inference in PyTorch with torch.compile + lit-llama code☆10Updated 2 years ago
- ☆21Updated 10 months ago
- Experimental GPU language with meta-programming☆24Updated last year
- ☆135Updated last year
- "PyTorch in Rust"☆17Updated last year
- An implementation of the Llama architecture, to instruct and delight☆21Updated 7 months ago
- Experimental compiler for deep learning models☆74Updated 4 months ago
- GPU based FFT written in Rust and CubeCL☆28Updated 3 weeks ago
- ☆32Updated last week
- Automatically derive Python dunder methods for your Rust code☆20Updated 9 months ago
- Experiment of using Tangent to autodiff triton☆81Updated last year
- Cute layout visualization☆27Updated this week
- Sample Python extension using Rust/PyO3/tch to interact with PyTorch☆40Updated last year
- A high-performance constrained decoding engine based on context free grammar in Rust☆58Updated 7 months ago
- ☆19Updated last month
- ☆91Updated last year
- 👷 Build compute kernels☆213Updated this week
- Profile your CoreML models directly from Python 🐍☆29Updated 4 months ago
- ☆39Updated last year