Skip to content
#

amd-strix-halo

Here are 4 public repositories matching this topic...

Language: All
Filter by language

Self-hosted, browser-based AI dungeon RPG, fully local on an AMD Strix Halo APU: narrator + per-character agents on one local LLM (Gemma 4 26B MoE via llama.cpp/Vulkan), local images (FLUX.2 klein/ComfyUI), expressive voice (Maya1 TTS). FastAPI brain, SQLite, vanilla JS, docker-compose with guided setup (CLI wizard or double-click HTML).

  • Updated Jul 25, 2026
  • Python

Model-agnostic NPU+GPU+CPU inference engine for AMD Strix Halo. FastFlowLM reverse-engineered — 19 architectures, 46+ 1BP models, 5 backends. GGUF/ONNX/Q4NX/1BP. Zero Python. MIT.

  • Updated Aug 5, 2026
  • C++

Improve this page

Add a description, image, and links to the amd-strix-halo topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the amd-strix-halo topic, visit your repo's landing page and select "manage topics."

Learn more