Open-source self-hosted home AI inference platform for AMD Strix Halo — multi-backend slots, OpenAI-compatible gateway, Vue 3 + FastAPI + systemd.
☆72Sep 6, 2026Updated this week
Alternatives and similar repositories for hal0
Users that are interested in hal0 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- vLLM + Qwen3.6-27B (BF16) OpenAI-compatible inference server on AMD Strix Halo (Ryzen AI Max+ 395, gfx1151). Vision input, 256K context, …☆17Apr 26, 2026Updated 4 months ago
- ☆26Jul 30, 2026Updated last month
- Home-enthusiast's guide to fine-tuning 27B+ LLMs on AMD Strix Halo (gfx1151, Ryzen AI MAX+ 395) — the patches and tuning to make Linux ma…☆27Jun 9, 2026Updated 3 months ago
- ☆108Aug 28, 2026Updated 2 weeks ago
- NEW ROCmfp4 format for llama.cpp☆158Jun 13, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- LLM inference in C/C++☆156Updated this week
- This tool helps you easily deploy ASR models on NPUs on AMD's Ryzen AI 300 series laptops☆27Jun 30, 2026Updated 2 months ago
- Read-Only source code mirror, Proxmox uses mailing list workflow for development.☆27Aug 21, 2026Updated 3 weeks ago
- ROCmFPX Family for AMD Hardware and Processors. More quants and special agent quants☆386Aug 22, 2026Updated 3 weeks ago
- ☆20Jul 21, 2026Updated last month
- Experimental support for many TTS/STT LLMs wrapped in a Wyoming API for consumption via Homeassistant☆43Aug 25, 2026Updated 2 weeks ago
- A converter for transferring gguf Q4_0, Q4_1 to FLM Q4NX☆39Aug 27, 2026Updated 2 weeks ago
- ☆253Oct 30, 2025Updated 10 months ago
- Fresh builds of llama.cpp with AMD ROCm™ 7 acceleration☆730Updated this week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- This repository builds llama.cpp for Strix Halo devices.☆69Updated this week
- Pure C wrapper library to use llama.cpp with Linux and Windows as simple as possible.☆15Sep 6, 2026Updated last week
- Watch all Kubernetes Resources☆16Jun 16, 2026Updated 2 months ago
- ☆518Sep 2, 2026Updated last week
- Use Lemonade LLM server with VS Code GitHub Copilot Chat☆23Jun 23, 2026Updated 2 months ago
- Prompts and model configs used in my videos.☆253Updated this week
- Marlin firmware for the Tevo Black Widow☆13Dec 6, 2017Updated 8 years ago
- ☆61Aug 17, 2026Updated 3 weeks ago
- This hook integrates with Claude Code to automatically create git checkpoints (snapshots) of your code before any file modifications. It …☆15Jul 6, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Collection of official scripts created by the Dione Team.☆15Feb 21, 2026Updated 6 months ago
- Tensor library for machine learning☆35Jul 31, 2026Updated last month
- Bootstrap GIT repo for setting up a Kodi repository☆13Aug 30, 2022Updated 4 years ago
- Give text-only LLMs vision. A tiny OpenAI-compatible proxy that lets reasoning models (DeepSeek, Qwen, GLM…) see images by querying a sep…☆49Jul 7, 2026Updated 2 months ago
- this is an easy way to make ai podcast useing ai loccaly like ollama and the tts of piper☆16Feb 17, 2026Updated 6 months ago
- Linux kernel driver for AMD Zen CPU monitoring (Zen 1-5): temperature, voltage, current, and power via SVI2/RAPL. Multi-file architecture…☆52Jan 8, 2026Updated 8 months ago
- One-command installer for integrating Google NotebookLM with OpenClaw via MCP.☆37Apr 18, 2026Updated 4 months ago
- pi.dev + llama.cpp ❤️☆25Jul 6, 2026Updated 2 months ago
- A reference of hardware requirements for Gen AI models☆15Apr 30, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Repository of Docker builds for Oracle databases.☆18Jul 3, 2023Updated 3 years ago
- ☆18Dec 1, 2025Updated 9 months ago
- Local LLM Server Manager + LlaMA.cpp + Chat☆135Updated this week
- Specialized fork for (relatively) fast single-GPU inference (in CUDA) using large MoE models that don't fit fully into VRAM☆17May 6, 2026Updated 4 months ago
- UI-based Fine-Tuning for Large Language Models (LLMs)☆20Dec 4, 2025Updated 9 months ago
- AI agents running research on single-GPU nanochat training automatically☆70Mar 16, 2026Updated 5 months ago
- ☆35Jan 14, 2026Updated 7 months ago