That's cool, but the real question is, why aren't you using Hipfire[1] or HaloPFX[2]? Both are far superior to llama.cpp in terms of performance, for Strix Halo.
[1] https://github.com/warpfront/hipfire
[2] https://github.com/julianmb/halofpx