huggingface / candle-cublasltLinks
☆13Updated last year
Alternatives and similar repositories for candle-cublaslt
Users that are interested in candle-cublaslt are comparing it to the libraries listed below
Sorting:
- Proof of concept for running moshi/hibiki using webrtc☆19Updated 3 months ago
- A small python library to run iterators in a separate process☆10Updated last year
- Rust crate for some audio utilities☆23Updated 2 months ago
- GPU based FFT written in Rust and CubeCL☆22Updated 2 months ago
- implement llava using candle☆15Updated 11 months ago
- 👷 Build compute kernels☆44Updated this week
- Read and write tensorboard data using Rust☆21Updated last year
- ☆11Updated 4 months ago
- Unleash the full potential of exascale LLMs on consumer-class GPUs, proven by extensive benchmarks, with no long-term adjustments and min…☆25Updated 6 months ago
- ☆19Updated 7 months ago
- Tensor library for Zig☆11Updated 6 months ago
- ESRGAN implemented in rust with candle☆16Updated last year
- Graph model execution API for Candle☆13Updated 6 months ago
- Implementing the BitNet model in Rust☆37Updated last year
- 8-bit floating point types for Rust☆45Updated 2 months ago
- Modular Rust transformer/LLM library using Candle☆35Updated last year
- The official evaluation suite and dynamic data release for MixEval.☆11Updated 8 months ago
- ☆12Updated last year
- A small rust-based data loader☆24Updated 5 months ago
- Inference Llama 2 in one file of zero-dependency, zero-unsafe Rust☆38Updated last year
- Nexusflow function call, tool use, and agent benchmarks.☆19Updated 5 months ago
- Training hybrid models for dummies.☆21Updated 4 months ago
- An extension library to Candle that provides PyTorch functions not currently available in Candle☆39Updated last year
- A text embedding extension for the Polars Dataframe library.☆24Updated 6 months ago
- Transformers provides a simple, intuitive interface for Rust developers who want to work with Large Language Models locally, powered by t…☆16Updated this week
- Generate Glue Code in seconds to simplify your Nvidia Triton Inference Server Deployments☆20Updated 11 months ago
- ☆39Updated 2 years ago
- JAX bindings for the flash-attention3 kernels☆11Updated 9 months ago
- Rust bindings for CTranslate2☆14Updated last year
- A high-performance constrained decoding engine based on context free grammar in Rust☆52Updated last week