Random AI notes for working with local models or playing around with random machine learning bits.
☆63Jun 7, 2026Updated 2 months ago
Alternatives and similar repositories for ai-notes
Users that are interested in ai-notes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆25Jul 30, 2026Updated last month
- ☆20Jul 21, 2026Updated last month
- vLLM Qwen 3.6-27B (AWQ-INT4) + DFlash speculative decoding on AMD Strix Halo (gfx1151 iGPU, 128 GB UMA, ROCm 7.13). 24.8 t/s single-strea…☆51May 10, 2026Updated 3 months ago
- ☆511Updated this week
- This is a mirror of the Strix Halo HomeLab wiki, to browse the wiki click on the link below☆73Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- LLM inference in C/C++☆17Jul 29, 2026Updated last month
- ☆150Updated this week
- Portable vLLM builds with AMD ROCm acceleration for Lemonade☆70Updated this week
- Open-source self-hosted home AI inference platform for AMD Strix Halo — multi-backend slots, OpenAI-compatible gateway, Vue 3 + FastAPI +…☆69Updated this week
- Profile-guided GPU kernel optimizer for AMD/RDNA3. Auto-tunes llama.cpp MMVQ kernels per model shape. 2x decode speedup on 7900 XTX.☆66Jul 7, 2026Updated last month
- ☆1,902Updated this week
- ☆66Updated this week
- ☆28Jun 10, 2026Updated 2 months ago
- vLLM + Qwen3.6-27B (BF16) OpenAI-compatible inference server on AMD Strix Halo (Ryzen AI Max+ 395, gfx1151). Vision input, 256K context, …☆17Apr 26, 2026Updated 4 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- LLM Fine Tuning Toolbox images for Ryzen AI 395+ Strix Halo☆66Sep 12, 2025Updated 11 months ago
- ☆251Oct 30, 2025Updated 10 months ago
- Fresh builds of llama.cpp with AMD ROCm™ 7 acceleration☆707Updated this week
- GPU monitor for Linux terminal supporting single or multiple gpu's in realtime☆21Jul 19, 2026Updated last month
- ☆21Sep 4, 2025Updated 11 months ago
- This repository builds llama.cpp for Strix Halo devices.☆68Updated this week
- AMD APU compatible Ollama. Get up and running with OpenAI gpt-oss, DeepSeek-R1, Gemma 3 and other models.☆16Aug 18, 2026Updated 2 weeks ago
- Dockernized ComfyUI with PyTorch & flash-attention for gfx1151 (AMD Strix Halo, Ryzen AI Max+ 395), relying on AMD's pre-built and pre-co…☆46Feb 25, 2026Updated 6 months ago
- vLLM for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆31Aug 25, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆106Updated this week
- Run LLMs on AMD Ryzen™ AI NPUs in minutes; purpose-built and deeply optimized for the AMD NPUs.☆1,834Updated this week
- Curated AMD GPU compatibility index and CLI for AI workloads☆31May 23, 2026Updated 3 months ago
- Linux driver for the embedded controller on the Sixunited AXB35-02 board.☆68Aug 20, 2026Updated 2 weeks ago
- A high-throughput and memory-efficient inference and serving engine for LLMs - Optimized for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI…☆86Jun 23, 2026Updated 2 months ago
- An opinionated, "speed" and "usability" focused agentic TUI with a built-in MCP registry/plugin system.☆33Apr 14, 2026Updated 4 months ago
- RDNA-native LLM inference engine in Rust.☆591Updated this week
- A test docker setup for running Qwen Code on StrixHalo☆17Feb 12, 2026Updated 6 months ago
- Proxy for OpenAI☆16Sep 2, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Evidence-backed AMD Strix Halo local-AI setup and benchmarks: Qwen3.8, Ollama, llama.cpp, Vulkan/ROCm, large GGUFs, and cross-OEM results…☆314Updated this week
- World's first AMD NPU driver for TUXEDO laptops - Enable AI acceleration on Linux☆51Aug 5, 2025Updated last year
- WikiPBX is an open source PBX web interface for FreeSWITCH. WikiPBX is written in python and uses the Django web application framework. C…☆10May 30, 2014Updated 12 years ago
- Proof of concept: Exploiting temporal coherence in LLM inference-- delta encoding for KV cache compression and weight-skip prediction. …☆50Apr 10, 2026Updated 4 months ago
- Local AI setup for AMD Strix Halo APU - Lemonade + Vulkan + kyuz0☆73Updated this week
- ComfyUI with Flash Attention for AI+ MAX 395 (gfx1151)☆20Jan 12, 2026Updated 7 months ago
- The HIP Environment and ROCm Kit - A lightweight open source build system for HIP and ROCm☆1,270Updated this week