AMD Strix Halo / Ryzen AI Halo local LLM setup and benchmark guide for Ryzen AI MAX+ 395 and Radeon 8060S: Ollama, llama.cpp Vulkan/RADV, ROCm, 101 t/s Qwen3-Coder, CHADROCK MTP, 120B GGUF, and raw evidence.
☆228Jul 18, 2026Updated this week
Alternatives and similar repositories for strix-halo-guide
Users that are interested in strix-halo-guide are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆55Jun 8, 2026Updated last month
- vLLM Qwen 3.6-27B (AWQ-INT4) + DFlash speculative decoding on AMD Strix Halo (gfx1151 iGPU, 128 GB UMA, ROCm 7.13). 24.8 t/s single-strea…☆47May 10, 2026Updated 2 months ago
- NEW ROCmfp4 format for llama.cpp☆135Jun 13, 2026Updated last month
- Home-enthusiast's guide to fine-tuning 27B+ LLMs on AMD Strix Halo (gfx1151, Ryzen AI MAX+ 395) — the patches and tuning to make Linux ma…☆26Jun 9, 2026Updated last month
- ☆1,774Jul 12, 2026Updated last week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Local AI setup for AMD Strix Halo APU - Lemonade + Vulkan + kyuz0☆52Updated this week
- Portable vLLM builds with AMD ROCm acceleration for Lemonade☆62Jul 10, 2026Updated last week
- Configuration and documentation for optimizing Ubuntu 24.04 on AMD Ryzen AI Max+ 395 with Radeon 8060S for LLM inference using llama.cpp …☆80Jul 9, 2026Updated last week
- This is a mirror of the Strix Halo HomeLab wiki, to browse the wiki click on the link below☆68Updated this week
- Linux driver for the embedded controller on the Sixunited AXB35-02 board.☆58Jul 11, 2026Updated last week
- ☆468Jun 17, 2026Updated last month
- Open-source self-hosted home AI inference platform for AMD Strix Halo — multi-backend slots, OpenAI-compatible gateway, Vue 3 + FastAPI +…☆56Updated this week
- ROCmFPX Family for AMD Hardware and Processors. More quants and special agent quants☆131Updated this week
- ☆141May 29, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Fresh builds of llama.cpp with AMD ROCm™ 7 acceleration☆623Updated this week
- LLM Fine Tuning Toolbox images for Ryzen AI 395+ Strix Halo☆63Sep 12, 2025Updated 10 months ago
- ☆56Updated this week
- LLM inference in C/C++☆83Updated this week
- ☆21Mar 13, 2026Updated 4 months ago
- Random AI notes for working with local models or playing around with random machine learning bits.☆60Jun 7, 2026Updated last month
- ☆23Mar 21, 2026Updated 4 months ago
- Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https…☆5,014Updated this week
- ☆73Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆248Oct 30, 2025Updated 8 months ago
- The HIP Environment and ROCm Kit - A lightweight open source build system for HIP and ROCm☆1,160Updated this week
- Official RepRapFirmware Configuration Tool☆14Jun 29, 2026Updated 3 weeks ago
- llama.cpp fork with AMD XDNA2 NPU backend for Ryzen AI MAX (npu5/XDNA2) — matrix multiply offload via XRT☆18Apr 1, 2026Updated 3 months ago
- ☆16Updated this week
- RDNA-native LLM inference engine in Rust.☆486Updated this week
- Decentralizing distribution of open-source AI models.☆16Updated this week
- ☆96Mar 8, 2026Updated 4 months ago
- Dockernized ComfyUI with PyTorch & flash-attention for gfx1151 (AMD Strix Halo, Ryzen AI Max+ 395), relying on AMD's pre-built and pre-co…☆40Feb 25, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Curated AMD GPU compatibility index and CLI for AI workloads☆31May 23, 2026Updated last month
- Run LLMs on AMD Ryzen™ AI NPUs in minutes. Just like Ollama - but purpose-built and deeply optimized for the AMD NPUs.☆1,636Updated this week
- This guide is beginner-friendly, project-driven, and laser-focused on the commands & concepts you will actually use while working with Do…☆16Dec 20, 2025Updated 7 months ago
- Local LLM Server Manager + LlaMA.cpp + Chat☆17Updated this week
- Use Lemonade LLM server with VS Code GitHub Copilot Chat☆20Jun 23, 2026Updated 3 weeks ago
- DelugeProbe☆13Oct 22, 2025Updated 9 months ago
- Collection of official scripts created by the Dione Team.☆15Feb 21, 2026Updated 5 months ago