huggingface / candle-cublasltLinks
☆12Updated last year
Alternatives and similar repositories for candle-cublaslt
Users that are interested in candle-cublaslt are comparing it to the libraries listed below
Sorting:
- A collection of optimisers for use with candle☆45Updated last week
- implement llava using candle☆15Updated last year
- Dataflow is a data processing library, primarily for machine learning.☆24Updated 2 years ago
- GPU based FFT written in Rust and CubeCL☆28Updated last week
- Unleash the full potential of exascale LLMs on consumer-class GPUs, proven by extensive benchmarks, with no long-term adjustments and min…☆26Updated last year
- ☆13Updated 2 weeks ago
- A single-binary, GPU-accelerated LLM server (HTTP and WebSocket API) written in Rust☆79Updated last year
- Rust crate for some audio utilities☆25Updated 10 months ago
- Rust Implementation of micrograd☆53Updated last year
- Fast and versatile tokenizer for language models, compatible with SentencePiece, Tokenizers, Tiktoken and more. Supports BPE, Unigram and…☆40Updated 2 months ago
- Rust crates for XetHub☆75Updated last year
- A simple, CUDA or CPU powered, library for creating vector embeddings using Candle and models from Hugging Face☆46Updated last year
- Rust SDK and CLI for Swarm Framework with Multi-Agent Orchestration☆15Updated 8 months ago
- Build tools for LLMs in Rust using Model Context Protocol☆38Updated 10 months ago
- Read and write tensorboard data using Rust☆24Updated last year
- Modular Rust transformer/LLM library using Candle☆37Updated last year
- Inference Llama 2 in one file of zero-dependency, zero-unsafe Rust☆39Updated 2 years ago
- ☆140Updated last year
- 8-bit floating point types for Rust☆62Updated last month
- This library supports evaluating disparities in generated image quality, diversity, and consistency between geographic regions.☆20Updated last year
- A collection of serverless apps that show how Fermyon's Serverless AI (currently in private beta) works. Reference: https://developer.fer…☆49Updated last year
- A high-performance constrained decoding engine based on context free grammar in Rust☆57Updated 7 months ago
- Fast serverless LLM inference, in Rust.☆108Updated 2 months ago
- This repository has code for fine-tuning LLMs with GRPO specifically for Rust Programming using cargo as feedback☆114Updated 10 months ago
- Proof of concept for running moshi/hibiki using webrtc☆19Updated 10 months ago
- Implementing the BitNet model in Rust☆43Updated last year
- ESRGAN implemented in rust with candle☆17Updated 2 years ago
- Experimental compiler for deep learning models☆73Updated 3 months ago
- ☆19Updated last week
- webgpu autograd library☆33Updated 7 months ago