Technical docs to help you make you Halo Strix WORK!
☆71Jan 10, 2026Updated 7 months ago
Alternatives and similar repositories for AMD-Strix-Halo-AI-Guide
Users that are interested in AMD-Strix-Halo-AI-Guide are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆147Aug 11, 2026Updated last week
- Open-source self-hosted home AI inference platform for AMD Strix Halo — multi-backend slots, OpenAI-compatible gateway, Vue 3 + FastAPI +…☆67Updated this week
- ☆17Jul 21, 2026Updated 3 weeks ago
- nixos config for amd strix-halo devices☆23Updated this week
- Dockernized ComfyUI with PyTorch & flash-attention for gfx1151 (AMD Strix Halo, Ryzen AI Max+ 395), relying on AMD's pre-built and pre-co…☆46Feb 25, 2026Updated 5 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆1,864Updated this week
- Strix Halo guide for AMD Ryzen AI MAX+ 395 / Radeon 8060S local LLM setup and benchmarks: Ollama, llama.cpp, Vulkan/RADV, ROCm, GGUF, and…☆280Updated this week
- ROCm/AMD GPU benchmark suite for llama.cpp, whisper.cpp, PyTorch☆21Feb 4, 2026Updated 6 months ago
- vLLM + Qwen3.6-27B (BF16) OpenAI-compatible inference server on AMD Strix Halo (Ryzen AI Max+ 395, gfx1151). Vision input, 256K context, …☆17Apr 26, 2026Updated 3 months ago
- This is a mirror of the Strix Halo HomeLab wiki, to browse the wiki click on the link below☆69Updated this week
- This repository builds llama.cpp for Strix Halo devices.☆63Updated this week
- Hierarchical RAG architecture scaling to 693K chunks on consumer hardware (4GB VRAM). Features 3-address routing, hybrid vector+graph fus…☆39Feb 11, 2026Updated 6 months ago
- Local inference with GMKTek evo-x2☆26Nov 3, 2025Updated 9 months ago
- Local AI setup for AMD Strix Halo APU - Lemonade + Vulkan + kyuz0☆66Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆56Jul 31, 2026Updated 2 weeks ago
- ☆175Apr 7, 2026Updated 4 months ago
- Portable vLLM builds with AMD ROCm acceleration for Lemonade☆67Updated this week
- VellumForge2 is a Golang CLI for generating high-quality Direct Preference Optimization datasets via a hierarchical prompt pipeline with …☆21Feb 15, 2026Updated 6 months ago
- UI-based Fine-Tuning for Large Language Models (LLMs)☆20Dec 4, 2025Updated 8 months ago
- The HIP Environment and ROCm Kit - A lightweight open source build system for HIP and ROCm☆1,210Updated this week
- NEW ROCmfp4 format for llama.cpp☆146Jun 13, 2026Updated 2 months ago
- Fresh builds of llama.cpp with AMD ROCm™ 7 acceleration☆666Updated this week
- ☆21Jul 4, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆502Updated this week
- Fork of turbo quant tom-tom and TurboQuant KV cache from domwox☆30Jul 7, 2026Updated last month
- Tensor library for machine learning☆36Jul 31, 2026Updated 2 weeks ago
- Watch all Kubernetes Resources☆16Jun 16, 2026Updated 2 months ago
- 30 tok/s for 20B MoE on 8 GB VRAM. Flat throughput to 32K context. Native MXFP4 + GGUF Q4_K/Q5_K/Q6_K via ggml CUDA kernels — zero dequan…☆21Apr 7, 2026Updated 4 months ago
- ☆11May 5, 2020Updated 6 years ago
- A test docker setup for running Qwen Code on StrixHalo☆17Feb 12, 2026Updated 6 months ago
- ☆250Oct 30, 2025Updated 9 months ago
- Official repository of AMD Playbooks☆98Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Puppet Guide☆12Jan 2, 2022Updated 4 years ago
- vLLM Qwen 3.6-27B (AWQ-INT4) + DFlash speculative decoding on AMD Strix Halo (gfx1151 iGPU, 128 GB UMA, ROCm 7.13). 24.8 t/s single-strea…☆51May 10, 2026Updated 3 months ago
- Ansible Role - Oracle WebLogic Server☆12Nov 27, 2020Updated 5 years ago
- Generic card scanner meant for card games with similar card formats as MTG & Pokemon. Takes an image, scans for a card, scans card for na…☆14Sep 7, 2020Updated 5 years ago
- Random AI notes for working with local models or playing around with random machine learning bits.☆63Jun 7, 2026Updated 2 months ago
- FlashQLA TileLang GDN kernels ported to NVIDIA Blackwell consumer (GB10 / DGX Spark)☆17Jun 5, 2026Updated 2 months ago
- User Script for http://workflowy.com that adds some extra features.☆12Oct 22, 2017Updated 8 years ago