Skip to content
#

strix-halo

Here are 75 public repositories matching this topic...

strix-halo-guide

AMD Strix Halo / Ryzen AI Halo local LLM setup and benchmark guide for Ryzen AI MAX+ 395 and Radeon 8060S: Ollama, llama.cpp Vulkan/RADV, ROCm, 101 t/s Qwen3-Coder, CHADROCK MTP, 120B GGUF, and raw evidence.

  • Updated Jul 21, 2026
  • Python

llama.cpp OpenAI-compatible server on Vulkan for AMD Strix Halo (gfx1151), GGUF weights pinned to GTT not VRAM. Serves poolside Laguna S 2.1, Gemma 4 and Qwen3.6 GGUFs on the stock Vulkan image, plus an opt-in ROCmFP4 + MTP stack (Ubuntu 26.04 + TheRock ROCm 7.13). Docker Compose, with real measured benchmarks.

  • Updated Jul 22, 2026
  • Python

Improve this page

Add a description, image, and links to the strix-halo topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the strix-halo topic, visit your repo's landing page and select "manage topics."

Learn more