Skip to content
#

strix-halo

Here are 73 public repositories matching this topic...

strix-halo-guide

AMD Strix Halo / Ryzen AI Halo local LLM setup and benchmark guide for Ryzen AI MAX+ 395 and Radeon 8060S: Ollama, llama.cpp Vulkan/RADV, ROCm, 101 t/s Qwen3-Coder, CHADROCK MTP, 120B GGUF, and raw evidence.

  • Updated Jul 18, 2026
  • Python

llama.cpp OpenAI-compatible server on Vulkan for AMD Strix Halo (gfx1151), GGUF weights pinned to GTT not VRAM. Stock-image Vulkan stack plus an opt-in ROCmFP4 + MTP stack: Ubuntu 26.04 + TheRock ROCm 7.13, built for plunderstruck's Qwen3.6 quants. Docker Compose, with real measured benchmarks.

  • Updated Jul 13, 2026
  • Python

Improve this page

Add a description, image, and links to the strix-halo topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the strix-halo topic, visit your repo's landing page and select "manage topics."

Learn more