Run LLMs on AMD Ryzen™ AI NPUs in minutes; purpose-built and deeply optimized for the AMD NPUs.
☆1,821Aug 28, 2026Updated this week
Alternatives and similar repositories for FastFlowLM
Users that are interested in FastFlowLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https…☆5,527Updated this week
- Build AI agents for your PC☆1,533Updated this week
- Close-to-metal programming for AMD NPUs☆135Updated this week
- AMD XDNA Driver (amdxdna.ko) for Linux and XRT SHIM library☆612Updated this week
- AMD Ryzen™ AI Software includes the tools and runtime libraries for optimizing and deploying AI inference on AMD Ryzen™ AI powered PCs.☆877Aug 18, 2026Updated last week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 🐉 Revolutionary NPU framework for Linux | 24,988 FPS face recognition | AMD XDNA support | World's first complete NPU stack☆43Aug 7, 2025Updated last year
- Fresh builds of llama.cpp with AMD ROCm™ 7 acceleration☆696Updated this week
- ☆1,887Updated this week
- A close-to-metal Python API for programming AMD Ryzen™ AI NPUs (AI Engines), built on an open-source MLIR-based compiler toolchain.☆682Updated this week
- Flm Companion is a lightweight Windows desktop app that simplifies managing your local FastFlowLM CLI server. It offers an intuitive inte…☆22Jul 21, 2026Updated last month
- ☆48Updated this week
- ☆508Aug 17, 2026Updated last week
- World's first AMD NPU driver for TUXEDO laptops - Enable AI acceleration on Linux☆51Aug 5, 2025Updated last year
- This tool helps you easily deploy ASR models on NPUs on AMD's Ryzen AI 300 series laptops☆27Jun 30, 2026Updated 2 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- The HIP Environment and ROCm Kit - A lightweight open source build system for HIP and ROCm☆1,257Updated this week
- ☆251Oct 30, 2025Updated 10 months ago
- Random AI notes for working with local models or playing around with random machine learning bits.☆63Jun 7, 2026Updated 2 months ago
- Tensor library for machine learning☆36Jul 31, 2026Updated 3 weeks ago
- ☆25Jul 30, 2026Updated last month
- Fork of LLVM to support AMD AIEngine processors☆210Updated this week
- Fedora COPR packaging for FastFlowLM - Run LLMs on AMD Ryzen AI NPUs☆16Updated this week
- Run LLMs on AMD Ryzen AI NPU (Linux)☆28Mar 29, 2026Updated 5 months ago
- This is a mirror of the Strix Halo HomeLab wiki, to browse the wiki click on the link below☆72Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- An open source branch of AIE API☆19Jun 12, 2026Updated 2 months ago
- vLLM Qwen 3.6-27B (AWQ-INT4) + DFlash speculative decoding on AMD Strix Halo (gfx1151 iGPU, 128 GB UMA, ROCm 7.13). 24.8 t/s single-strea…☆51May 10, 2026Updated 3 months ago
- ☆23Updated this week
- RDNA-native LLM inference engine in Rust.☆563Updated this week
- LLM Fine Tuning Toolbox images for Ryzen AI 395+ Strix Halo☆66Sep 12, 2025Updated 11 months ago
- AMD APU compatible Ollama. Get up and running with OpenAI gpt-oss, DeepSeek-R1, Gemma 3 and other models.☆161May 15, 2026Updated 3 months ago
- Let's have some retro gaming fun with AI! Join the discord: https://discord.gg/5xXzkMu8Zk☆99Nov 19, 2025Updated 9 months ago
- No-code CLI designed for accelerating ONNX workflows☆245Jul 1, 2026Updated last month
- ROCm Library Files for gfx1103 and update with others arches based on AMD GPUs for use in Windows.☆889May 18, 2026Updated 3 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Community maintained mobile app for use with Lemonade. Join our discord: https://discord.gg/5xXzkMu8Zk☆50Jul 28, 2026Updated last month
- vLLM + Qwen3.6-27B (BF16) OpenAI-compatible inference server on AMD Strix Halo (Ryzen AI Max+ 395, gfx1151). Vision input, 256K context, …☆17Apr 26, 2026Updated 4 months ago
- ☆150Aug 11, 2026Updated 2 weeks ago
- ☆21Sep 4, 2025Updated 11 months ago
- llama.cpp fork with additional SOTA quants and improved performance☆3,144Updated this week
- llama.cpp fork with AMD XDNA2 NPU backend for Ryzen AI MAX (npu5/XDNA2) — matrix multiply offload via XRT☆20Apr 1, 2026Updated 4 months ago
- NPUEval is an LLM evaluation dataset written specifically to target AIE kernel code generation on RyzenAI hardware.☆32Nov 8, 2025Updated 9 months ago