☆61Sep 14, 2026Updated 3 weeks ago
Alternatives and similar repositories for vllm-openvino
Users that are interested in vllm-openvino are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Run Generative AI models with simple C++/Python API and using OpenVINO Runtime☆600Updated this week
- Community maintained hardware plugin for vLLM on Intel Gaudi☆61Updated this week
- With OpenVINO Test Drive, users can run large language models (LLMs) and models trained by Intel Geti on their devices, including AI PCs …☆41Sep 1, 2026Updated last month
- SGLang kernel library for Intel XPU☆36Updated this week
- OpenVINO Intel NPU Compiler☆102Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A curated list of OpenVINO based AI projects☆212May 18, 2026Updated 4 months ago
- This repository contains Dockerfiles, scripts, yaml files, Helm charts, etc. used to scale out AI containers with versions of TensorFlow …☆79Sep 8, 2026Updated last month
- 🤗 Optimum Intel: Accelerate inference with Intel optimization tools☆622Updated this week
- OpenVINO Post-Training Optimization Toolkit Tutorial☆17Sep 28, 2020Updated 6 years ago
- OpenAI Triton backend for Intel® GPUs☆273Updated this week
- Run cpplint with reviewdog☆13Oct 2, 2026Updated last week
- Add genai backend for ollama to run generative AI models using OpenVINO Runtime.☆31Apr 16, 2026Updated 5 months ago
- ☆15Apr 11, 2024Updated 2 years ago
- Developer kits reference setup scripts for various kinds of Intel platforms and GPUs☆54Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆115Updated this week
- Benchmark Suite Invocation Scripting☆11Mar 16, 2022Updated 4 years ago
- ☆15May 8, 2025Updated last year
- Repository for OpenVINO's extra modules☆187Updated this week
- DLL注入工具☆13Nov 9, 2020Updated 5 years ago
- matmul using AMX instructions☆24May 7, 2024Updated 2 years ago
- SPDK fork of nvme-cli. No longer supported - use standard nvme-cli with SPDK nvme CUSE instead. See https://spdk.io/doc/nvme.html#nvme_…☆15Apr 10, 2024Updated 2 years ago
- Deep Learning Inference benchmark. Supports OpenVINO™ toolkit, TensorFlow, TensorFlow Lite, ONNX Runtime, OpenCV DNN, MXNet, PyTorch, Apa…☆36Aug 10, 2026Updated 2 months ago
- ONNX Runtime: cross-platform, high performance scoring engine for ML models☆92Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- GPU Functional Descriptor for memory access☆35May 24, 2026Updated 4 months ago
- How to export PyTorch models with unsupported layers to ONNX and then to Intel OpenVINO☆28Feb 20, 2025Updated last year
- OpenGL 学习代码☆15Jun 25, 2023Updated 3 years ago
- An Awesome list of oneAPI projects☆166Aug 8, 2025Updated last year
- Pre-built components and code samples to help you build and deploy production-grade AI applications with the OpenVINO™ Toolkit from Intel☆217Jul 30, 2026Updated 2 months ago
- Llama causal LM fully recreated in LibTorch. Designed to be used in Unreal Engine 5☆16Sep 19, 2024Updated 2 years ago
- ☆15Jun 26, 2024Updated 2 years ago
- Meta project around MLIR☆58Updated this week
- ☆197Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆11Jan 7, 2023Updated 3 years ago
- OpenVINO™ is an open source toolkit for optimizing and deploying AI inference☆10,974Updated this week
- Software kit for Qualcomm Cloud AI 100☆19Sep 3, 2026Updated last month
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.☆14Jan 8, 2026Updated 9 months ago
- Hexagon-MLIR is a compiler toolchain for compiling and executing AI kernels and models on Qualcomm Hexagon Neural Processing Units (NPUs)…☆257Oct 3, 2026Updated last week
- code repo for GCR [FAST'26]☆17Mar 3, 2026Updated 7 months ago
- PilotFish harvests the free GPU cycles of cloud gaming with deep learning training☆14Jul 2, 2022Updated 4 years ago