The ggml‑org team released version b11322 of llama.cpp on Oct 1 2026. The update bundles pre‑compiled binaries for Apple Silicon macOS, Intel macOS, iOS, and multiple Ubuntu builds (CPU, Vulkan, CUDA). A race‑condition fix in the hex‑workqueue component is also included. The binaries are available as direct download links on the GitHub release page.
Why it matters
Developers can now run llama.cpp locally on more platforms without building from source, speeding up testing and deployment.