huggingface / candle-cublasltLinks
☆12Updated last year
Alternatives and similar repositories for candle-cublaslt
Users that are interested in candle-cublaslt are comparing it to the libraries listed below
Sorting:
- Unleash the full potential of exascale LLMs on consumer-class GPUs, proven by extensive benchmarks, with no long-term adjustments and min…☆25Updated 11 months ago
- A collection of optimisers for use with candle☆43Updated 2 months ago
- implement llava using candle☆15Updated last year
- Rust SDK and CLI for Swarm Framework with Multi-Agent Orchestration☆15Updated 6 months ago
- Implementing the BitNet model in Rust☆40Updated last year
- A single-binary, GPU-accelerated LLM server (HTTP and WebSocket API) written in Rust☆79Updated last year
- Chunk Dedupe Estimation☆20Updated 11 months ago
- Rust crates for XetHub☆70Updated last year
- A complete(grpc service and lib) Rust inference with multilingual embedding support. This version leverages the power of Rust for both GR…☆39Updated last year
- Rust Implementation of micrograd☆53Updated last year
- First token cutoff sampling inference example☆30Updated last year
- ☆139Updated last year
- Fast and versatile tokenizer for language models, compatible with SentencePiece, Tokenizers, Tiktoken and more. Supports BPE, Unigram and…☆36Updated 2 weeks ago
- ☆134Updated last year
- ☆40Updated this week
- This repository has code for fine-tuning LLMs with GRPO specifically for Rust Programming using cargo as feedback☆108Updated 7 months ago
- A high-performance constrained decoding engine based on context free grammar in Rust☆55Updated 5 months ago
- ☆12Updated 9 months ago
- Read and write tensorboard data using Rust☆23Updated last year
- Dataflow is a data processing library, primarily for machine learning.☆24Updated 2 years ago
- Framework-Agnostic RL Environments for LLM Fine-Tuning☆37Updated last week
- Modular Rust transformer/LLM library using Candle☆36Updated last year
- Fast serverless LLM inference, in Rust.☆94Updated 7 months ago
- Inference Llama 2 in one file of zero-dependency, zero-unsafe Rust☆39Updated 2 years ago
- ☆39Updated 3 years ago
- A collection of serverless apps that show how Fermyon's Serverless AI (currently in private beta) works. Reference: https://developer.fer…☆50Updated 10 months ago
- The official evaluation suite and dynamic data release for MixEval.☆11Updated last year
- 8-bit floating point types for Rust☆60Updated 2 months ago
- Because it's there.☆16Updated last year
- IBM development fork of https://github.com/huggingface/text-generation-inference☆61Updated last month