Evidence-backed AMD Strix Halo local-AI setup and benchmarks: Qwen3.8, Ollama, llama.cpp, Vulkan/ROCm, large GGUFs, and cross-OEM results.
☆336Sep 19, 2026Updated this week
Alternatives and similar repositories for strix-halo-guide
Users that are interested in strix-halo-guide are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆56Sep 1, 2026Updated 2 weeks ago
- vLLM Qwen 3.6-27B (AWQ-INT4) + DFlash speculative decoding on AMD Strix Halo (gfx1151 iGPU, 128 GB UMA, ROCm 7.13). 24.8 t/s single-strea…☆52May 10, 2026Updated 4 months ago
- NEW ROCmfp4 format for llama.cpp☆158Jun 13, 2026Updated 3 months ago
- ☆1,944Updated this week
- Local AI setup for AMD Strix Halo APU - Lemonade + Vulkan + kyuz0☆82Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Portable vLLM builds with AMD ROCm acceleration for Lemonade☆76Sep 9, 2026Updated last week
- Technical docs to help you make you Halo Strix WORK!☆71Sep 3, 2026Updated 2 weeks ago
- Configuration and documentation for optimizing Ubuntu 24.04 on AMD Ryzen AI Max+ 395 with Radeon 8060S for LLM inference using llama.cpp …☆101Aug 19, 2026Updated last month
- This is a mirror of the Strix Halo HomeLab wiki, to browse the wiki click on the link below☆73Aug 31, 2026Updated 2 weeks ago
- Linux driver for the embedded controller on the Sixunited AXB35-02 board.☆75Aug 20, 2026Updated last month
- ☆522Updated this week
- ☆29Jun 17, 2026Updated 3 months ago
- Fresh builds of llama.cpp with AMD ROCm™ 7 acceleration☆741Sep 11, 2026Updated last week
- ROCmFPX Family for AMD Hardware and Processors. More quants and special agent quants☆397Aug 22, 2026Updated 3 weeks ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆61Aug 17, 2026Updated last month
- This repository builds llama.cpp for Strix Halo devices.☆69Updated this week
- Random AI notes for working with local models or playing around with random machine learning bits.☆63Jun 7, 2026Updated 3 months ago
- ☆27Jul 30, 2026Updated last month
- Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https…☆5,752Updated this week
- ☆255Oct 30, 2025Updated 10 months ago
- The HIP Environment and ROCm Kit - A lightweight open source build system for HIP and ROCm☆1,324Updated this week
- llama.cpp fork with AMD XDNA2 NPU backend for Ryzen AI MAX (npu5/XDNA2) — matrix multiply offload via XRT☆21Apr 1, 2026Updated 5 months ago
- ☆20Jul 21, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- RDNA-native LLM inference engine in Rust.☆634Updated this week
- Decentralizing distribution of open-source AI models.☆24Updated this week
- Profile-guided GPU kernel optimizer for AMD/RDNA3. Auto-tunes llama.cpp MMVQ kernels per model shape. 2x decode speedup on 7900 XTX.☆69Sep 5, 2026Updated 2 weeks ago
- Curated AMD GPU compatibility index and CLI for AI workloads☆31May 23, 2026Updated 3 months ago
- Run LLMs on AMD Ryzen™ AI NPUs in minutes; purpose-built and deeply optimized for the AMD NPUs.☆1,884Updated this week
- A high-performance RCCL / NCCL (ROCm Communication Collectives Library) plugin for Thunderbolt 5 that enables GPU-to-GPU communication ac…☆68Updated this week
- LLM speculative inference server for heterogeneous hardware & consumer GPUs☆2,868Updated this week
- Collection of official scripts created by the Dione Team.☆15Feb 21, 2026Updated 6 months ago
- DelugeProbe☆13Oct 22, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Local LLM Server Manager + LlaMA.cpp + Chat☆140Sep 7, 2026Updated last week
- A TUI around llama.cpp for running, managing, and benchmarking local GGUF models and launching the pi coding agent against your local ser…☆31Updated this week
- Close-to-metal programming for AMD NPUs☆140Updated this week
- The Unified Model Registry for all your local AI apps.☆65Apr 9, 2026Updated 5 months ago
- A task based orchestration TUI for oh-my-pi agent harness enabling highly parallel agentic development.☆26Feb 21, 2026Updated 7 months ago
- This tool helps you easily deploy ASR models on NPUs on AMD's Ryzen AI 300 series laptops☆28Jun 30, 2026Updated 2 months ago
- Run Pi coding agent isolated in a Docker Sandbox microVM with a local llama-server as the inference backend☆28Jun 16, 2026Updated 3 months ago