Open-source bring-up + verified matmul on the first-gen AMD XDNA1 (Phoenix/Hawk Point) NPU on Linux — the gen FastFlowLM/Lemonade skip. RyzenAI-npu1, mlir-aie/IRON, XRT.
-
Updated
Jul 6, 2026 - Shell
Open-source bring-up + verified matmul on the first-gen AMD XDNA1 (Phoenix/Hawk Point) NPU on Linux — the gen FastFlowLM/Lemonade skip. RyzenAI-npu1, mlir-aie/IRON, XRT.
Model-agnostic NPU+GPU+CPU inference engine for AMD Strix Halo. FastFlowLM fully reverse-engineered and replaced with a native open-source XDNA 2 stack — zero proprietary code. GGUF/ONNX/1BP ternary (TQ2 2-bit). Zero Python at runtime. MIT.
Add a description, image, and links to the npu-inference topic page so that developers can more easily learn about it.
To associate your repository with the npu-inference topic, visit your repo's landing page and select "manage topics."