A Pure Rust based LLM, VLM, VLA, TTS, OCR Inference Engine, powering by Candle & Rust. Alternate to your llama.cpp but much more simpler and cleaner..
โ475Sep 20, 2026Updated this week
Alternatives and similar repositories for Crane
Users that are interested in Crane are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ๐ฅ๐ฅ Kokoro in Rust. https://huggingface.co/hexgrad/Kokoro-82M Insanely fast, realtime TTS with high quality you ever have.โ824Aug 4, 2026Updated last month
- Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.โ728Updated this week
- โ650Sep 14, 2026Updated last week
- Blazing-fast LLM inference in pure Rust. No PyTorch and Python runtime.โ325Sep 9, 2026Updated last week
- Rust standalone inference of Namo-500M series models. Extremly tiny, runing VLM on CPU.โ24Mar 12, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI โข AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Fast ML inference & training for ONNX models in Rustโ2,516Updated this week
- Kokoro TTS็Rustๆจ็ๅฎ็ฐโ35Jun 1, 2026Updated 3 months ago
- Rust implementation of Qwen3-TTS speech synthesisโ147Feb 6, 2026Updated 7 months ago
- Rust bindings for OpenNMT/CTranslate2โ60Aug 31, 2026Updated 2 weeks ago
- Fast, flexible LLM inferenceโ7,712Sep 8, 2026Updated last week
- Rust bindings to https://github.com/k2-fsa/sherpa-onnxโ314Mar 8, 2026Updated 6 months ago
- A wrapper around the llama-cpp library for rust, including new Sampler API from llama-cpp.โ49Updated this week
- aha model inference library, now supports Qwen(2.5VL/3/3VL/3.5/ASR/3Embedding/3Reranker), MiniCPM(4/5), VoxCPM(0.5B/1.5/2), DeepSeek-OCR/โฆโ394Jun 7, 2026Updated 3 months ago
- A Rust ๐ฆ port of the Hugging Face smolagents library.โ43Mar 26, 2025Updated last year
- End-to-end encrypted email - Proton Mail โข AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Rust library for generating vector embeddings and reranking locally!โ1,016Updated this week
- An educational Rust project for exporting and running inference on Qwen3 LLM familyโ44Aug 3, 2025Updated last year
- Instant, controllable, local pre-trained AI models in Rustโ2,230Updated this week
- Ultrafast PyTorch-like AI Framework Written from Ground-Up in Rustโ104Mar 18, 2026Updated 6 months ago
- โ18Jul 1, 2025Updated last year
- A native Rust toolkit for OCR, document layout analysis, and vision-language document understanding.โ172Updated this week
- very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rustโ396Updated this week
- Democratizing large model inference and training on any device.โ269Sep 4, 2026Updated 2 weeks ago
- Minimalist ML framework for Rustโ21,080Updated this week
- Deploy on Railway without the complexity - Free Credits Offer โข AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Unofficial Rust bindings to Apple's mlx frameworkโ373Updated this week
- โ27Jan 29, 2026Updated 7 months ago
- โ๏ธ๐ฆ Build modular and scalable LLM Applications in Rustโ8,681Updated this week
- Adding a multi-text multi-speaker script (diffe) that is based on a script from asiff00 on issue 61 for Sesame: A Conversational Speech Gโฆโ26Mar 28, 2025Updated last year
- โ24Jan 22, 2025Updated last year
- Fast string seaching for node.js ( aho-corasick algorithm )โ12Jan 24, 2024Updated 2 years ago
- Open-source LLM/VLM load balancer and serving platform for self-hosting LLMs (and VLMs) at scale ๐๐ฆ Alternative to projects like llm-d,โฆโ1,675Updated this week
- Fast serverless LLM inference, in Rust.โ133Sep 11, 2026Updated last week
- โ17Aug 30, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI โข AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Rust bindings for bitnet.cpp based on llama-cpp-4โ17Dec 28, 2025Updated 8 months ago
- RWKV-based Text-to-Speech implementation in Rustโ28Oct 14, 2025Updated 11 months ago
- Ridiculously fast & accurate streaming voice activity detectionโ201Aug 30, 2026Updated 3 weeks ago
- Models and examples built with Burnโ382Sep 8, 2026Updated last week
- A comprehensive Rust translation of the code from Sebastian Raschka's Build an LLM from Scratch book.โ335Updated this week
- โ15Mar 18, 2026Updated 6 months ago
- The official Rust SDK for the Model Context Protocolโ3,945Updated this week