Open-source bring-up + verified matmul on the first-gen AMD XDNA1 (Phoenix/Hawk Point) NPU on Linux — the gen FastFlowLM/Lemonade skip. RyzenAI-npu1, mlir-aie/IRON, XRT.
-
Updated
Aug 6, 2026 - Shell
Open-source bring-up + verified matmul on the first-gen AMD XDNA1 (Phoenix/Hawk Point) NPU on Linux — the gen FastFlowLM/Lemonade skip. RyzenAI-npu1, mlir-aie/IRON, XRT.
Model-agnostic, hardware-agnostic pure-C++ inference engine — one binary, NPU + GPU + CPU. 94% HF architecture coverage. Reverse-engineered AMD's closed-source Strix Halo NPU stack in 4 days. GGUF/ONNX/1BP. Zero Python. MIT.
Add a description, image, and links to the npu-inference topic page so that developers can more easily learn about it.
To associate your repository with the npu-inference topic, visit your repo's landing page and select "manage topics."