The llama.cpp project released version b11294 on 30 September 2026. The update adds row‑prefetching for Windows, meaning the library can load model data in advance to speed up generation. The change was contributed by Pranesh Gonegandla and merged without further modification. It also includes minor clean‑up and build fixes.
Why it matters
Windows users of llama.cpp will see faster model inference thanks to the new prefetch logic.