☆657Oct 9, 2026Updated this week
Alternatives and similar repositories for llama-cpp-rs
Users that are interested in llama-cpp-rs are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- High-level, optionally asynchronous Rust bindings to llama.cpp☆251Jun 5, 2024Updated 2 years ago
- LLama.cpp rust bindings☆424Jun 27, 2024Updated 2 years ago
- Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.☆732Sep 17, 2026Updated 3 weeks ago
- Rust library for generating vector embeddings and reranking locally!☆1,033Updated this week
- Fast ML inference & training for ONNX models in Rust☆2,554Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A wrapper around the llama-cpp library for rust, including new Sampler API from llama-cpp.☆51Updated this week
- A Pure Rust based LLM, VLM, VLA, TTS, OCR Inference Engine, powering by Candle & Rust. Alternate to your llama.cpp but much more simpler …☆498Updated this week
- The Easiest Rust Interface for Local LLMs and an Interface for Deterministic Signals from Probabilistic LLM Vibes☆254Aug 6, 2025Updated last year
- Rust bindings to https://github.com/k2-fsa/sherpa-onnx☆314Mar 8, 2026Updated 7 months ago
- Rust bindings to https://github.com/leejet/stable-diffusion.cpp☆55Sep 30, 2026Updated last week
- Unofficial Rust bindings to Apple's mlx framework☆379Oct 3, 2026Updated last week
- The official Rust SDK for the Model Context Protocol☆3,992Updated this week
- Fast, flexible LLM inference☆7,736Oct 1, 2026Updated last week
- Open-source LLM/VLM load balancer and serving platform for self-hosting LLMs (and VLMs) at scale 🏓🦙 Alternative to projects like llm-d,…☆1,675Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A simple and easy-to-use library for interacting with the Ollama API.☆1,062Updated this week
- pyannote audio diarization in rust☆129Sep 7, 2025Updated last year
- Minimalist ML framework for Rust☆21,150Updated this week
- Rust client for the huggingface hub aiming for minimal subset of features over `huggingface-hub` python package☆337Updated this week
- Democratizing large model inference and training on any device.☆267Sep 4, 2026Updated last month
- Rust bindings to https://github.com/ggerganov/whisper.cpp☆945Jul 30, 2025Updated last year
- Safer rust wrapper over mnn☆24Mar 5, 2026Updated 7 months ago
- Apple frameworks bindings for rust☆220Updated this week
- Speech detection using silero vad in Rust☆34Dec 16, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- 🦜️🔗LangChain for Rust, the easiest way to write LLM-based programs in Rust☆1,349Updated this week
- ⚙️🦀 Build modular and scalable LLM Applications in Rust☆8,843Updated this week
- Rust multiprovider generative AI client (Ollama, OpenAi, Anthropic, Gemini, DeepSeek, ZAI, OpenRouter, FireworksAI, xAI/Grok, Groq,, ...)☆900Updated this week
- Burn is a next generation tensor library and Deep Learning Framework that doesn't compromise on flexibility, efficiency and portability.☆16,090Updated this week
- Instant, controllable, local pre-trained AI models in Rust☆2,228Updated this week
- Fast, streaming indexing, query, and agentic LLM applications in Rust☆789Updated this week
- ONNX neural network inference engine☆341Updated this week
- A high-level idiomatic Rust wrapper around Pdfium, the C++ PDF library used by the Google Chromium project.☆713Aug 16, 2026Updated last month
- ☆13Nov 4, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Multi-platform high-performance compute language extension for Rust.☆2,425Updated this week
- A cross-platform inference engine for neural TTS models.☆75Nov 25, 2024Updated last year
- Simple, efficient and cross-platform TFIDF-based text summarizer in Rust☆12Apr 12, 2024Updated 2 years ago
- Rust bindings for the C++ api of PyTorch.☆5,498Aug 23, 2026Updated last month
- Rust native ready-to-use NLP pipelines and transformer-based models (BERT, DistilBERT, GPT2,...)☆3,077Jan 13, 2026Updated 8 months ago
- Tiny, no-nonsense, self-contained, Tensorflow and ONNX inference☆3,086Updated this week
- A cross-platform browser ML framework.☆773Sep 24, 2026Updated 2 weeks ago