huggingface / candle-cublasltLinks

☆12

Alternatives and similar repositories for candle-cublaslt

Users that are interested in candle-cublaslt are comparing it to the libraries listed below

Sorting:

kyegomez / Exa
Unleash the full potential of exascale LLMs on consumer-class GPUs, proven by extensive benchmarks, with no long-term adjustments and min…
☆25Updated last year
KGrewal1 / candle-optimisers
A collection of optimisers for use with candle
☆43Updated 3 months ago
eugenehp / gpu-fft
GPU based FFT written in Rust and CubeCL
☆24Updated 5 months ago
yaman / fashion-clip-rs
A complete(grpc service and lib) Rust inference with multilingual embedding support. This version leverages the power of Rust for both GR…
☆39Updated last year
EricLBuehler / candle_graphs
Graph model execution API for Candle
☆16Updated 3 months ago
chenwanqq / candle-llava
implement llava using candle
☆15Updated last year
georgesheth / swarms-rust
Rust SDK and CLI for Swarm Framework with Multi-Agent Orchestration
☆15Updated 7 months ago
Systemcluster / kitoken
Fast and versatile tokenizer for language models, compatible with SentencePiece, Tokenizers, Tiktoken and more. Supports BPE, Unigram and…
☆38Updated last month
Dan-wanna-M / kbnf
A high-performance constrained decoding engine based on context free grammar in Rust
☆55Updated 5 months ago
oxidized-transformers / oxidized-transformers
Modular Rust transformer/LLM library using Candle
☆36Updated last year
fermyon / ai-examples
A collection of serverless apps that show how Fermyon's Serverless AI (currently in private beta) works. Reference: https://developer.fer…
☆50Updated 11 months ago
pixelspark / poly
A single-binary, GPU-accelerated LLM server (HTTP and WebSocket API) written in Rust
☆79Updated last year
EricLBuehler / float8
8-bit floating point types for Rust
☆60Updated 3 months ago
jafioti / dataflow
Dataflow is a data processing library, primarily for machine learning.
☆24Updated 2 years ago
ShelbyJenkins / candle_embed
A simple, CUDA or CPU powered, library for creating vector embeddings using Candle and models from Hugging Face
☆46Updated last year
philschmid / MixEval
The official evaluation suite and dynamic data release for MixEval.
☆11Updated last year
dylibso / chainsocket
Proof of concept for a generative AI application framework powered by WebAssembly and Extism
☆14Updated 2 years ago
kanpuriyanawab / picograd
Rust Implementation of micrograd
☆53Updated last year
tomsanbear / bitnet-rs
Implementing the BitNet model in Rust
☆41Updated last year
Vaibhavs10 / fast-llm.rs
☆139Updated last year
octoml / triton-client-rs
A client library in Rust for Nvidia Triton.
☆30Updated 2 years ago
leo-du / llama2.rs
Inference Llama 2 in one file of zero-dependency, zero-unsafe Rust
☆39Updated 2 years ago
google-deepmind / multiscope
Real-time visualisation
☆20Updated last month
mokeyish / candle-ext
An extension library to Candle that provides PyTorch functions not currently available in Candle
☆40Updated last year
kyutai-labs / moshi-webrtc
Proof of concept for running moshi/hibiki using webrtc
☆19Updated 8 months ago
lancedb / tantivy-object-store
Tantivy directory implementation backed by object_store
☆37Updated last year
kyutai-labs / yomikomi
A small rust-based data loader
☆32Updated this week
LaurentMazare / tboard-rs
Read and write tensorboard data using Rust
☆23Updated last year
avahowell / offeryn
Build tools for LLMs in Rust using Model Context Protocol
☆38Updated 8 months ago
jbellis / colbert-live
ColBERT for live vector indexes
☆28Updated last year