aha model inference library, now supports Qwen(2.5VL/3/3VL/3.5/ASR/3Embedding/3Reranker), MiniCPM(4/5), VoxCPM(0.5B/1.5/2), DeepSeek-OCR/2, Hunyuan-OCR, PaddleOCR-VL/1.5/1.6, RMBG2.0, GLM(ASR-Nano-2512/OCR), Fun-ASR-Nano-2512, LFM(2/2.5/2VL/2.5VL)
☆383Jun 7, 2026Updated 2 months ago
Alternatives and similar repositories for aha
Users that are interested in aha are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆31Sep 1, 2025Updated 11 months ago
- ☆15Aug 25, 2025Updated 11 months ago
- Rust (using burn) implementation of VoxCPM☆22Apr 9, 2026Updated 4 months ago
- ☆21May 2, 2026Updated 3 months ago
- Bare-Metal Rust Audio AI Framework☆176Jun 13, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Tool for speech recognition using sensevoice-small☆32Jun 11, 2026Updated 2 months ago
- A Pure Rust based LLM, VLM, VLA, TTS, OCR Inference Engine, powering by Candle & Rust. Alternate to your llama.cpp but much more simpler …☆452Updated this week
- A high-performance Rust implementation of llama.cpp - LLM inference engine with full GGUF support☆18Apr 7, 2026Updated 4 months ago
- Standalone C++ inference project for VoxCPM models built on top of ggml.☆89Jul 14, 2026Updated last month
- Fast ML inference & training for ONNX models in Rust☆2,446Updated this week
- Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.☆712Aug 2, 2026Updated last week
- Reusable UI components and widgets built on top of gpui.☆341Jul 16, 2026Updated 3 weeks ago
- Blazing-fast LLM inference in pure Rust. No PyTorch and Python runtime.☆302Updated this week
- Implementation of Qwen3-ASR-0.6B in GGML☆108Jul 28, 2026Updated 2 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Video creation tool: Video editing (with extensive AI-assisted features), screen recording, streaming, and screen sharing. 视频编辑(大量AI辅助功能…☆54Updated this week
- A wrapper around the llama-cpp library for rust, including new Sampler API from llama-cpp.☆43Updated this week
- Rust bindings to https://github.com/k2-fsa/sherpa-onnx☆312Mar 8, 2026Updated 5 months ago
- Rust Agent Development Kit (ADK-Rust): Build AI agents in Rust with modular components for models, tools, memory, realtime voice, and mor…☆599Updated this week
- 高性能OCR识别库,支持上百种语言,提供命令行、图形界面及C API多种调用方式,使用便捷高效。 High-performance OCR library powered by PaddleOCR v4/v5/v6 with MNN backend. Supports 1…☆304Aug 5, 2026Updated last week
- ☆25Jan 29, 2026Updated 6 months ago
- 阿里云物联网设备端 Rust SDK / Alink 协议☆11Apr 23, 2023Updated 3 years ago
- Cross-platform abstractions for GPU compute☆49Jul 4, 2026Updated last month
- ☆50Mar 18, 2026Updated 4 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆631Updated this week
- Utilizes ONNX Runtime for TTS model.☆66Jul 30, 2026Updated 2 weeks ago
- GenUI is an innovative SFP front-end framework developed in the Rust language, was originally inspired by Vue3 and Makepad. Intended to h…☆89May 7, 2026Updated 3 months ago
- An ecosystem framework for Rust. IoC/Context/Config/Web/ORM☆51Aug 15, 2025Updated 11 months ago
- 将 Qwen3-ASR 的 LLM 部分导出为 GGUF,用 llama.cpp 进行加速推理。后者支持 Vulkan 和 Cuda 加速。☆212Apr 29, 2026Updated 3 months ago
- A Rust library integrated with ONNXRuntime, providing a collection of Computer Vison and Vision-Language models such as YOLO, FastVLM, an…☆439Jul 30, 2026Updated 2 weeks ago
- Build a Web Application Template for the Generalized Rust Axum Library☆32Jun 19, 2026Updated last month
- ☆25Jul 20, 2026Updated 3 weeks ago
- Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2☆643Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆35Aug 7, 2025Updated last year
- ☆13Nov 4, 2023Updated 2 years ago
- Rust library for generating vector embeddings, reranking locally!☆987Aug 4, 2026Updated last week
- An animation engine implemented in rust, inspired by manim and JAnim☆632Updated this week
- An educational Rust project for exporting and running inference on Qwen3 LLM family☆43Aug 3, 2025Updated last year
- 用 onnx 和 gguf 格式混合运行 Fun-ASR-Nano 模型全流程☆155May 5, 2026Updated 3 months ago
- The Rust SDK library supports the call of Hikvision network cameras and MVS industrial cameras☆14Dec 21, 2024Updated last year