The llama.cpp project added a build option for the ZDNN backend, a library that speeds up matrix math on IBM Z hardware. The change was committed by Aaron Teo on Sep 30 2026. The new code compiles the backend but does not run any tests yet.
Why it matters
It lets developers try llama.cpp on IBM Z systems, potentially improving performance for large‑language‑model inference.