☆30May 23, 2026Updated 2 months ago
Alternatives and similar repositories for llama-cpp-qnn-builder
Users that are interested in llama-cpp-qnn-builder are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LLM inference in C/C++☆53Jul 10, 2026Updated last week
- the original reference implementation of a specified llama.cpp backend for Qualcomm Hexagon NPU on Android phone, history of ggml-hexagon…☆48Updated this week
- snpe tutorial☆10Dec 25, 2023Updated 2 years ago
- Inference of YOLOv7 model applied on Qualcomm SNPE for pedestrian detection with embedded system.☆13Sep 23, 2024Updated last year
- A CUDA kernel for NHWC GroupNorm for PyTorch☆23Nov 15, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Hopenet: deep head pose estimator on ncnn☆10Jun 18, 2020Updated 6 years ago
- ☆11Feb 5, 2026Updated 5 months ago
- Ultra fast head pose estimation on a bare Raspberry Pi 4 at 20 FPS☆10Dec 21, 2021Updated 4 years ago
- High-speed and easy-use LLM serving framework for local deployment☆161Aug 7, 2025Updated 11 months ago
- High-performance system monitor for Rockchip SoCs (RK3588, RK3399) with real-time CPU, GPU, NPU, RGA, memory, and process monitoring. W…☆25Nov 25, 2025Updated 7 months ago
- Optimized pose detector inference for edge devices☆15Feb 23, 2023Updated 3 years ago
- FastRPC is Qualcomm's userspace library that facilitates efficient remote procedure calls between the CPU and DSP for high-performance co…☆104Jul 14, 2026Updated last week
- QAI AppBuilder is designed to help developers easily execute models on WoS and Linux platforms. It encapsulates the Qualcomm® AI Runtime …☆187Updated this week
- hexagon tutorial☆56Mar 29, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Shadowsocks/ShadowsocksR 账号在线监控☆12Nov 25, 2018Updated 7 years ago
- This repository is to share the EdgeAI Lab with Microcontrollers Series material to the entire community. We will share documents, presen…☆17Oct 14, 2021Updated 4 years ago
- Self-implemented NN operators for Qualcomm's Hexagon NPU☆76Sep 30, 2025Updated 9 months ago
- A Android Library for YOLOv5/YOLOv7/YOLOv8 Detection and Pose Inference Based on NCNN☆59Aug 15, 2024Updated last year
- deepstream + cuda,yolo26,yolo-master,yolo11,yolov8,sam,transformer, etc.☆27Feb 7, 2026Updated 5 months ago
- 智能家教微信小程序☆11Sep 15, 2018Updated 7 years ago
- End to End Speech to Speech with Emotion System☆15Feb 6, 2025Updated last year
- Tesseract tessdata downloader from GitHub repositories☆11Sep 17, 2021Updated 4 years ago
- 使用onnxruntime部署Gaze-LLE凝视目标估计,包含C++和Python两个版本的程序☆17Jan 21, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- pyTorch variational autoencoder, with explainations☆11May 31, 2017Updated 9 years ago
- ☆15Apr 28, 2023Updated 3 years ago
- 官方transformers源码解析。AI大模型时代,pytorch、transformer是新操作系统,其他都是运行在其上面的软件。☆16Sep 25, 2023Updated 2 years ago
- CUDA SGEMM optimization note☆15Oct 31, 2023Updated 2 years ago
- Bjontegaard metric calculation. Include BD-PSNR and BD-rate☆14Sep 4, 2024Updated last year
- An example app of DNNLibrary :)☆13Jul 26, 2019Updated 6 years ago
- ☆11May 19, 2025Updated last year
- PyCon mini 東海 2024 のトーク「Google Colaboratoryで試すVLM」で紹介したサンプル集☆12Nov 15, 2024Updated last year
- ☆13Apr 28, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Build environment for Raspberry Pi Pico (RP2040) C/C++ SDK☆11Jan 25, 2021Updated 5 years ago
- 🎵 Control YouTube players with browser by Alfred☆12Sep 30, 2020Updated 5 years ago
- [ACL2025 Oral🔥]Turning Trash into Treasure: Accelerating Inference of Large Language Models with Token Recycling☆29Nov 11, 2025Updated 8 months ago
- Code for paper "ElasticTrainer: Speeding Up On-Device Training with Runtime Elastic Tensor Selection" (MobiSys'23)☆14Nov 1, 2023Updated 2 years ago
- 本项目是一个通过文字生成图片的项目,基于开源模型Stable Diffusion V1.5生成可以在手机的CPU和NPU上运行的模型,包括其配套的模型运行框架。☆245Mar 29, 2024Updated 2 years ago
- ☆13Mar 18, 2024Updated 2 years ago
- RAG-QA is a free, containerised question-answer framework that allows you to ask questions to your documents in an intuitive way☆21Jan 25, 2024Updated 2 years ago