Hexagon-MLIR is a compiler toolchain for compiling and executing AI kernels and models on Qualcomm Hexagon Neural Processing Units (NPUs).
☆231Sep 16, 2026Updated this week
Alternatives and similar repositories for hexagon-mlir
Users that are interested in hexagon-mlir are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Self-implemented NN operators for Qualcomm's Hexagon NPU☆77Sep 30, 2025Updated 11 months ago
- ☆95Dec 16, 2025Updated 9 months ago
- Torq compiler sources☆56Jul 28, 2026Updated last month
- ☆11Sep 4, 2025Updated last year
- FastRPC is Qualcomm's userspace library that facilitates efficient remote procedure calls between the CPU and DSP for high-performance co…☆114Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Tenstorrent MLIR compiler☆306Updated this week
- IREE's PyTorch Frontend, based on Torch Dynamo.☆110Jul 20, 2026Updated last month
- the original FastRPC-based implementation of a specified llama.cpp backend for Qualcomm Hexagon NPU, history of ggml-hexagon: https://git…☆56Updated this week
- Fast Multimodal LLM on Mobile Devices☆1,613Sep 8, 2026Updated last week
- MLIR-based partitioning system☆210Updated this week
- MLIR-based toolkit targeting intel heterogeneous hardware☆54Jun 26, 2026Updated 2 months ago
- ☆17Updated this week
- onnxruntime-qnn is the Qualcomm AI Runtime (QAIRT) execution provider for onnxruntime. It provides onnxruntime hardware acceleration and …☆51Updated this week
- Repo for AI Compiler team. The intended purpose of this repo is for implementation of a PJRT device.☆74Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- High-speed and easy-use LLM serving framework for local deployment☆166Aug 7, 2025Updated last year
- A close-to-metal Python API for programming AMD Ryzen™ AI NPUs (AI Engines), built on an open-source MLIR-based compiler toolchain.☆689Updated this week
- Inference RWKV v5, v6 and v7 with Qualcomm AI Engine Direct SDK☆99Jul 27, 2026Updated last month
- ☆22Updated this week
- Wave: Python Domain-Specific Language for High Performance Machine Learning☆59Jun 29, 2026Updated 2 months ago
- Development repository for the Triton-Linalg conversion☆225Feb 7, 2025Updated last year
- A collection of out-of-tree extensions for the Triton language and compiler☆38Updated this week
- An MLIR frontend for tensor expressions☆24Sep 5, 2020Updated 6 years ago
- QAI AppBuilder is designed to help developers easily execute models on WoS and Linux platforms. It encapsulates the Qualcomm® AI Runtime …☆240Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Meta project around MLIR☆55Updated this week
- Code for "An Introduction to Tensor Tiling in MLIR" tutorial given at EuroLLVM 2025☆26Jun 5, 2025Updated last year
- A retargetable MLIR-based machine learning compiler and runtime toolkit.☆3,935Updated this week
- FlagTree is a unified compiler supporting multiple AI chip backends for custom Deep Learning operations, which is forked from triton-lang…☆352Updated this week
- TPP experimentation on MLIR for linear algebra☆165Updated this week
- Intel® Extension for MLIR. A staging ground for MLIR dialects and tools for Intel devices using the MLIR toolchain.☆156Updated this week
- OpenVINO Intel NPU Compiler☆100Updated this week
- Github mirror of trition-lang/triton repo.☆196Updated this week
- Representation and Reference Lowering of ONNX Models in MLIR Compiler Infrastructure☆1,056Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆11Oct 30, 2021Updated 4 years ago
- An Android kernel for kindle kt3☆13Jul 19, 2022Updated 4 years ago
- triton for dsa☆69Aug 25, 2026Updated 3 weeks ago
- Shared Middle-Layer for Triton Compilation☆34Aug 14, 2026Updated last month
- A Python compiler design toolkit.☆590Updated this week
- Embedded Universal DSL: a good DSL for us, by us☆79Updated this week
- A Python-embedded DSL that makes it easy to write fast, scalable ML kernels with minimal boilerplate.☆948Updated this week