QAI AppBuilder is designed to help developers easily execute models on WoS and Linux platforms. It encapsulates the Qualcomm® AI Runtime SDK APIs into a set of simplified interfaces for running models on the NPU/HTP.
☆240Sep 18, 2026Updated this week
Alternatives and similar repositories for qai-appbuilder
Users that are interested in qai-appbuilder are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Qualcomm® AI Hub apps are a collection of state-of-the-art machine learning models optimized for performance (latency, memory etc.) a…☆460Updated this week
- hexagon tutorial☆66Mar 29, 2026Updated 5 months ago
- ☆202Sep 5, 2026Updated 2 weeks ago
- Inference RWKV v5, v6 and v7 with Qualcomm AI Engine Direct SDK☆99Jul 27, 2026Updated last month
- Let's use Qualcomm NPU in Android☆21Feb 18, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [EMNLP Findings 2024] MobileQuant: Mobile-friendly Quantization for On-device Language Models☆69Sep 22, 2024Updated last year
- FastRPC is Qualcomm's userspace library that facilitates efficient remote procedure calls between the CPU and DSP for high-performance co…☆114Updated this week
- Stable Diffusion+LCM在SG2300X上,纵享丝滑一秒出图☆17Nov 29, 2024Updated last year
- 本项目是一个通过文字生成图片的项目,基于开源模型Stable Diffusion V1.5生成可以在手机的CPU和NPU上运行的模型,包括其配套的模型运行框架。☆248Mar 29, 2024Updated 2 years ago
- the original FastRPC-based implementation of a specified llama.cpp backend for Qualcomm Hexagon NPU, history of ggml-hexagon: https://git…☆56Updated this week
- some hexagon intrinsic examples based on Qualcomm Hexagon☆18Mar 7, 2025Updated last year
- Fast Multimodal LLM on Mobile Devices☆1,613Sep 8, 2026Updated last week
- Run Chinese MobileBert model on SNPE.☆15May 19, 2023Updated 3 years ago
- Hexagon-MLIR is a compiler toolchain for compiling and executing AI kernels and models on Qualcomm Hexagon Neural Processing Units (NPUs)…☆231Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- High-speed and easy-use LLM serving framework for local deployment☆166Aug 7, 2025Updated last year
- The open-source project for "Mandheling: Mixed-Precision On-Device DNN Training with DSP Offloading"[MobiCom'2022]☆20Aug 4, 2022Updated 4 years ago
- ☆346Feb 12, 2026Updated 7 months ago
- ☆29Jun 30, 2025Updated last year
- LLM inference in C/C++☆54Sep 3, 2026Updated 2 weeks ago
- snpe tutorial☆10Dec 25, 2023Updated 2 years ago
- A simple tutorial of SNPE.☆186Mar 30, 2023Updated 3 years ago
- This library empowers users to seamlessly port pretrained models and checkpoints on the HuggingFace (HF) hub (developed using HF transfor…☆99Updated this week
- C++ implementations for various tokenizers (sentencepiece, tiktoken etc).☆50Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Example apps and demos using PyTorch's ExecuTorch framework☆83Updated this week
- On-device AI across mobile, embedded and edge for PyTorch☆5,036Updated this week
- Large Language Model Onnx Inference Framework☆35Nov 25, 2025Updated 9 months ago
- ggml学习笔记,ggml是一个机器学习的推理框架☆18Mar 24, 2024Updated 2 years ago
- ☆17Updated this week
- Quantize yolov5 using pytorch_quantization.🚀🚀🚀☆15Oct 24, 2023Updated 2 years ago
- The repository supports TensorRT, QNN platform inference, 2D obstacle detection yolo series (yolov5, yolov8, yolo11, yolox), semantic seg…☆20May 6, 2025Updated last year
- MobileSAM のエンコーダー/デコーダーをONNXに変換し、推論するサンプル☆12Apr 11, 2024Updated 2 years ago
- 搜藏的希望的代码片段☆13Jun 6, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆23Nov 6, 2025Updated 10 months ago
- A lightweight, single-header C++11 Jinja2 template engine for LLM chat templates.☆20Mar 4, 2026Updated 6 months ago
- View Synthesis Function☆13Jun 6, 2016Updated 10 years ago
- ffmpeg+cuvid+tensorrt+multicamera☆12Dec 31, 2024Updated last year
- Learning Matchable Image Transformations☆13Sep 10, 2019Updated 7 years ago
- Text2speech & tone color conversion demo running on SG2300x 结合openvoice和emotivoice的TTS+即时克隆☆22Oct 30, 2024Updated last year
- stable diffusion using mnn☆68Sep 28, 2023Updated 2 years ago