A Pure Rust based LLM, VLM, VLA, TTS, OCR Inference Engine, powering by Candle & Rust. Alternate to your llama.cpp but much more simpler and cleaner..
β498Oct 10, 2026Updated this week
Alternatives and similar repositories for Crane
Users that are interested in Crane are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π₯π₯ Kokoro in Rust. https://huggingface.co/hexgrad/Kokoro-82M Insanely fast, realtime TTS with high quality you ever have.β827Updated this week
- Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.β732Sep 17, 2026Updated 3 weeks ago
- β657Updated this week
- Blazing-fast LLM inference in pure Rust. No PyTorch and Python runtime.β334Sep 9, 2026Updated last month
- Rust standalone inference of Namo-500M series models. Extremly tiny, runing VLM on CPU.β24Mar 12, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Fast ML inference & training for ONNX models in Rustβ2,554Updated this week
- Rust implementation of Qwen3-TTS speech synthesisβ148Feb 6, 2026Updated 8 months ago
- Fast, flexible LLM inferenceβ7,736Oct 1, 2026Updated last week
- Rust bindings for OpenNMT/CTranslate2β61Aug 31, 2026Updated last month
- Rust bindings to https://github.com/k2-fsa/sherpa-onnxβ314Mar 8, 2026Updated 7 months ago
- A wrapper around the llama-cpp library for rust, including new Sampler API from llama-cpp.β51Updated this week
- aha model inference library, now supports Qwen(2.5VL/3/3VL/3.5/ASR/3Embedding/3Reranker), MiniCPM(4/5), VoxCPM(0.5B/1.5/2), DeepSeek-OCR/β¦β392Jun 7, 2026Updated 4 months ago
- A Rust π¦ port of the Hugging Face smolagents library.β43Mar 26, 2025Updated last year
- Rust library for generating vector embeddings and reranking locally!β1,033Updated this week
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- An educational Rust project for exporting and running inference on Qwen3 LLM familyβ44Aug 3, 2025Updated last year
- Instant, controllable, local pre-trained AI models in Rustβ2,228Updated this week
- Ultrafast PyTorch-like AI Framework Written from Ground-Up in Rustβ104Mar 18, 2026Updated 6 months ago
- β18Jul 1, 2025Updated last year
- A native Rust toolkit for OCR, document layout analysis, and vision-language document understanding.β179Updated this week
- very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rustβ408Updated this week
- A Rust library integrated with ONNXRuntime, providing a collection of Computer Vison and Vision-Language models such as YOLO, FastVLM, anβ¦β444Jul 30, 2026Updated 2 months ago
- Democratizing large model inference and training on any device.β267Sep 4, 2026Updated last month
- Minimalist ML framework for Rustβ21,150Updated this week
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Unofficial Rust bindings to Apple's mlx frameworkβ379Oct 3, 2026Updated last week
- β27Jan 29, 2026Updated 8 months ago
- βοΈ π¦ Build modular and scalable LLM Applications in Rustβ8,843Updated this week
- Adding a multi-text multi-speaker script (diffe) that is based on a script from asiff00 on issue 61 for Sesame: A Conversational Speech Gβ¦β26Mar 28, 2025Updated last year
- β24Jan 22, 2025Updated last year
- Fast string seaching for node.js ( aho-corasick algorithm )β12Jan 24, 2024Updated 2 years ago
- Open-source LLM/VLM load balancer and serving platform for self-hosting LLMs (and VLMs) at scale ππ¦ Alternative to projects like llm-d,β¦β1,675Updated this week
- Fast serverless LLM inference, in Rust.β135Sep 11, 2026Updated last month
- Rust bindings for bitnet.cpp based on llama-cpp-4β17Dec 28, 2025Updated 9 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- High-performance Qwen3-TTS implementation | Instruction-driven Β· Zero-shot voice cloning Β· Streaming Β· RTF 0.55β70Jun 26, 2026Updated 3 months ago
- RWKV-based Text-to-Speech implementation in Rustβ28Oct 14, 2025Updated 11 months ago
- Ridiculously fast & accurate streaming voice activity detectionβ208Aug 30, 2026Updated last month
- Models and examples built with Burnβ384Oct 1, 2026Updated last week
- A comprehensive Rust translation of the code from Sebastian Raschka's Build an LLM from Scratch book.β337Updated this week
- A high-performance inference engine for AI modelsβ2,141Updated this week
- Utilities for encoding and decoding frames using `async/await`.β34Jul 9, 2025Updated last year