Intel® NPU Acceleration Library
☆713Apr 24, 2025Updated last year
Alternatives and similar repositories for intel-npu-acceleration-library
Users that are interested in intel-npu-acceleration-library are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Intel® NPU (Neural Processing Unit) Driver☆466Updated this week
- OpenVINO Intel NPU Compiler☆102Updated this week
- Library for modelling performance costs of different Neural Network workloads on NPU devices☆36Aug 31, 2026Updated last month
- Run Generative AI models with simple C++/Python API and using OpenVINO Runtime☆600Updated this week
- A Python package for extending the official PyTorch that can easily obtain performance on Intel platform☆2,011Mar 30, 2026Updated 6 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆18Oct 2, 2026Updated last week
- OpenAI Triton backend for Intel® GPUs☆273Updated this week
- OpenVINO LLM Benchmark☆11Dec 7, 2023Updated 2 years ago
- ☆20Nov 27, 2025Updated 10 months ago
- OpenVINO Tokenizers extension☆57Updated this week
- Accelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V,…☆8,849Jan 28, 2026Updated 8 months ago
- 🤗 Optimum Intel: Accelerate inference with Intel optimization tools☆622Updated this week
- OpenVINO™ is an open source toolkit for optimizing and deploying AI inference☆10,974Updated this week
- ☆61Dec 18, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Fork of LLVM to support AMD AIEngine processors☆212Oct 2, 2026Updated last week
- AMD Ryzen™ AI Software includes the tools and runtime libraries for optimizing and deploying AI inference on AMD Ryzen™ AI powered PCs.☆889Aug 18, 2026Updated last month
- oneAPI Level Zero Specification Headers and Loader☆337Updated this week
- ☆717Updated this week
- ⚡ Build your chatbot within minutes on your favorite device; offer SOTA compression techniques for LLMs; run LLMs efficiently on Intel Pl…☆2,167Oct 8, 2024Updated 2 years ago
- SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, …☆2,713Updated this week
- ☆292Updated this week
- Intel® Graphics Compute Runtime for oneAPI Level Zero and OpenCL™ Driver☆1,453Updated this week
- ☆154Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Chisel implementation of Neural Processing Unit for System on the Chip☆37Updated this week
- ☆266Apr 8, 2024Updated 2 years ago
- ☆13May 11, 2023Updated 3 years ago
- ☆61Sep 14, 2026Updated 3 weeks ago
- Open-source library of optimized deep learning operations (matmul, convolution, attention) for CPUs (x64, AArch64, RISC-V) and Intel GPUs…☆4,059Updated this week
- An innovative library for efficient LLM inference via low-bit quantization☆351Aug 30, 2024Updated 2 years ago
- ☆115Updated this week
- Olive: Simplify ML Model Finetuning, Conversion, Quantization, and Optimization for CPUs, GPUs and NPUs.☆2,397Updated this week
- A Gradio Web UI for running local LLM on Intel GPU (e.g., local PC with iGPU, discrete GPU such as Arc, Flex and Max) using IPEX-LLM.☆17Updated this week
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A close-to-metal Python API for programming AMD Ryzen™ AI NPUs (AI Engines), built on an open-source MLIR-based compiler toolchain.☆700Updated this week
- Tenstorrent MLIR compiler☆314Updated this week
- 📚 Jupyter notebook tutorials for OpenVINO™☆3,223Updated this week
- ☆25Sep 19, 2025Updated last year
- Development repository for the Triton language and compiler☆20,337Updated this week
- Lightweight Python Wrapper for OpenVINO, enabling LLM inference on NPUs☆30Dec 17, 2024Updated last year
- Generative AI extensions for onnxruntime☆1,135Updated this week