The llama.cpp project published release b11464 on 7 Oct 2026. The update fixes a SYCL (a standard for heterogeneous computing) bug that prevented mixed GPU models from running together. The change is in the file ggml‑sycl.cpp and was co‑authored by Georgi Gerganov. The release also lists many build targets, from macOS Apple Silicon to Linux CUDA 13.4 and Android Snapdragon platforms.
Why it matters
Developers can now run Llama.cpp on systems with different GPUs without errors, expanding hardware compatibility.