The ggml‑org team published llama.cpp version b11501 on Oct 8 2026. The update adds SYCL (a cross‑platform compute API) fast Walsh‑Hadamard transform optimizations, which can speed up certain matrix ops. It also ships ready‑to‑run binaries for macOS Apple Silicon, macOS Intel, iOS, and Ubuntu builds with CPU, Vulkan and CUDA 12.8 support.
Why it matters
Developers can now run Llama models faster on SYCL‑compatible GPUs and avoid building the library themselves.