A two-day, non-expert guide to running local LLMs on an AMD Strix Halo (Ryzen AI Max+ 395) box — full ~120GB memory pool, ROCm backend, NPU in parallel via FastFlowLM, one OpenAI-compatible endpoint.
-
Updated
Jul 25, 2026
A two-day, non-expert guide to running local LLMs on an AMD Strix Halo (Ryzen AI Max+ 395) box — full ~120GB memory pool, ROCm backend, NPU in parallel via FastFlowLM, one OpenAI-compatible endpoint.
The Operating System for Autonomous Intelligence on AMD AI Fabric — AMD Developer Hackathon ACT II
Add a description, image, and links to the amd-ai topic page so that developers can more easily learn about it.
To associate your repository with the amd-ai topic, visit your repo's landing page and select "manage topics."