QAI AppBuilder is designed to help developers easily execute models on WoS and Linux platforms. It encapsulates the Qualcomm® AI Runtime SDK APIs into a set of simplified interfaces for running models on the NPU/HTP.
☆185Jul 20, 2026Updated this week
Alternatives and similar repositories for qai-appbuilder
Users that are interested in qai-appbuilder are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- hexagon tutorial☆54Mar 29, 2026Updated 3 months ago
- Let's use Qualcomm NPU in Android☆21Feb 18, 2025Updated last year
- ☆199Updated this week
- onnxruntime-qnn is the Qualcomm AI Runtime (QAIRT) execution provider for onnxruntime. It provides onnxruntime hardware acceleration and …☆39Updated this week
- ☆29May 23, 2026Updated last month
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- the original reference implementation of a specified llama.cpp backend for Qualcomm Hexagon NPU on Android phone, history of ggml-hexagon…☆48Updated this week
- 本项目是一个通过文字生成图片的项目,基于开源模型Stable Diffusion V1.5生成可以在手机的CPU和NPU上运行的模型,包括其配套的模型运行框架。☆243Mar 29, 2024Updated 2 years ago
- some hexagon intrinsic examples based on Qualcomm Hexagon☆17Mar 7, 2025Updated last year
- Self-implemented NN operators for Qualcomm's Hexagon NPU☆75Sep 30, 2025Updated 9 months ago
- Fast Multimodal LLM on Mobile Devices☆1,572Jun 26, 2026Updated 3 weeks ago
- ☆10Jul 18, 2024Updated 2 years ago
- Hexagon-MLIR is a compiler toolchain for compiling and executing AI kernels and models on Qualcomm Hexagon Neural Processing Units (NPUs)…☆177Jul 2, 2026Updated 2 weeks ago
- This repository is a read-only mirror of https://gitlab.arm.com/kleidi/kleidiai☆171Updated this week
- Run Chinese MobileBert model on SNPE.☆15May 19, 2023Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- High-speed and easy-use LLM serving framework for local deployment☆161Aug 7, 2025Updated 11 months ago
- The open-source project for "Mandheling: Mixed-Precision On-Device DNN Training with DSP Offloading"[MobiCom'2022]☆20Aug 4, 2022Updated 3 years ago
- Demonstration of running a native LLM on Android device.☆257Updated this week
- LLM inference in C/C++☆53Jul 10, 2026Updated last week
- snpe tutorial☆10Dec 25, 2023Updated 2 years ago
- A simple tutorial of SNPE.☆186Mar 30, 2023Updated 3 years ago
- C++ implementations for various tokenizers (sentencepiece, tiktoken etc).☆50Updated this week
- Example apps and demos using PyTorch's ExecuTorch framework☆81Updated this week
- On-device AI across mobile, embedded and edge for PyTorch☆4,813Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Large Language Model Onnx Inference Framework☆35Nov 25, 2025Updated 7 months ago
- ggml学习笔记,ggml是一个机器学习的推理框架☆18Mar 24, 2024Updated 2 years ago
- Quantize yolov5 using pytorch_quantization.🚀🚀🚀☆15Oct 24, 2023Updated 2 years ago
- The repository supports TensorRT, QNN platform inference, 2D obstacle detection yolo series (yolov5, yolov8, yolo11, yolox), semantic seg…☆20May 6, 2025Updated last year
- MobileSAM のエンコーダー/デコーダーをONNXに変換し、推論するサンプル☆12Apr 11, 2024Updated 2 years ago
- This project is intended to build and deploy an SNPE model on Qualcomm Devices, which are having unsupported layers which are not part of…☆10Oct 4, 2021Updated 4 years ago
- 搜藏的希望的代码片段☆13Jun 6, 2023Updated 3 years ago
- A lightweight, single-header C++11 Jinja2 template engine for LLM chat templates.☆20Mar 4, 2026Updated 4 months ago
- ffmpeg+cuvid+tensorrt+multicamera☆12Dec 31, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ROCK 5A☆10Nov 21, 2025Updated 8 months ago
- Support PyTorch model conversion with LiteRT.☆1,069Updated this week
- Mobile (i.e., Android, iOS) foundation model (i.e., LLM, VLM) deployed with MLC☆26Feb 12, 2025Updated last year
- ☆72Feb 27, 2023Updated 3 years ago
- RKNN模型推理部署模板☆24Aug 11, 2023Updated 2 years ago
- ☆46Jun 30, 2026Updated 3 weeks ago
- Project is intended to build and deploy an scene detection application onto Qualcomm Robotics development Kit (RB5) that detects whether …☆10Jun 26, 2022Updated 4 years ago