A Pure Rust based LLM, VLM, VLA, TTS, OCR Inference Engine, powering by Candle & Rust. Alternate to your llama.cpp but much more simpler and cleaner..
โ460Aug 31, 2026Updated this week
Alternatives and similar repositories for Crane
Users that are interested in Crane are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ๐ฅ๐ฅ Kokoro in Rust. https://huggingface.co/hexgrad/Kokoro-82M Insanely fast, realtime TTS with high quality you ever have.โ816Aug 4, 2026Updated 3 weeks ago
- Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.โ718Aug 20, 2026Updated last week
- โ639Updated this week
- Blazing-fast LLM inference in pure Rust. No PyTorch and Python runtime.โ309Updated this week
- Rust standalone inference of Namo-500M series models. Extremly tiny, runing VLM on CPU.โ24Mar 12, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available โข AdRun AI, ML, and HPC workloads on powerful cloud GPUsโwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Fast ML inference & training for ONNX models in Rustโ2,479Updated this week
- Kokoro TTS็Rustๆจ็ๅฎ็ฐโ36Jun 1, 2026Updated 2 months ago
- Rust implementation of Qwen3-TTS speech synthesisโ147Feb 6, 2026Updated 6 months ago
- Rust bindings for OpenNMT/CTranslate2โ59Aug 5, 2026Updated 3 weeks ago
- Fast, flexible LLM inferenceโ7,639Updated this week
- Rust bindings to https://github.com/k2-fsa/sherpa-onnxโ311Mar 8, 2026Updated 5 months ago
- A wrapper around the llama-cpp library for rust, including new Sampler API from llama-cpp.โ46Updated this week
- aha model inference library, now supports Qwen(2.5VL/3/3VL/3.5/ASR/3Embedding/3Reranker), MiniCPM(4/5), VoxCPM(0.5B/1.5/2), DeepSeek-OCR/โฆโ390Jun 7, 2026Updated 2 months ago
- A Rust ๐ฆ port of the Hugging Face smolagents library.โ44Mar 26, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI โข AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Rust library for generating vector embeddings and reranking locally!โ1,001Updated this week
- An educational Rust project for exporting and running inference on Qwen3 LLM familyโ44Aug 3, 2025Updated last year
- Instant, controllable, local pre-trained AI models in Rustโ2,225Aug 22, 2026Updated last week
- Ultrafast PyTorch-like AI Framework Written from Ground-Up in Rustโ103Mar 18, 2026Updated 5 months ago
- โ18Jul 1, 2025Updated last year
- A native Rust toolkit for OCR, document layout analysis, and vision-language document understanding.โ161Aug 20, 2026Updated last week
- very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rustโ390Jul 28, 2026Updated last month
- A Rust library integrated with ONNXRuntime, providing a collection of Computer Vison and Vision-Language models such as YOLO, FastVLM, anโฆโ441Jul 30, 2026Updated last month
- Democratizing large model inference and training on any device.โ268Aug 21, 2026Updated last week
- Virtual machines for every use case on DigitalOcean โข AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Minimalist ML framework for Rustโ20,977Updated this week
- Unofficial Rust bindings to Apple's mlx frameworkโ370Updated this week
- โ26Jan 29, 2026Updated 7 months ago
- โ๏ธ๐ฆ Build modular and scalable LLM Applications in Rustโ8,469Updated this week
- Adding a multi-text multi-speaker script (diffe) that is based on a script from asiff00 on issue 61 for Sesame: A Conversational Speech Gโฆโ26Mar 28, 2025Updated last year
- โ24Jan 22, 2025Updated last year
- Fast string seaching for node.js ( aho-corasick algorithm )โ12Jan 24, 2024Updated 2 years ago
- Open-source LLM/VLM load balancer and serving platform for self-hosting LLMs (and VLMs) at scale ๐๐ฆ Alternative to projects like llm-d,โฆโ1,665Jul 19, 2026Updated last month
- โ17Aug 30, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient โข AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Rust bindings for bitnet.cpp based on llama-cpp-4โ17Dec 28, 2025Updated 8 months ago
- Fast serverless LLM inference, in Rust.โ130Updated this week
- High-performance Qwen3-TTS implementation | Instruction-driven ยท Zero-shot voice cloning ยท Streaming ยท RTF 0.55โ67Jun 26, 2026Updated 2 months ago
- RWKV-based Text-to-Speech implementation in Rustโ28Oct 14, 2025Updated 10 months ago
- Ridiculously fast & accurate streaming voice activity detectionโ197Updated this week
- Models and examples built with Burnโ379Aug 12, 2026Updated 2 weeks ago
- A comprehensive Rust translation of the code from Sebastian Raschka's Build an LLM from Scratch book.โ333Updated this week