The open‑source llama.cpp library released version b11323 on Oct 1 2026. It adds compiled binaries for macOS Apple Silicon, Intel, iOS, and multiple Ubuntu configurations, including CPU, Vulkan and CUDA builds. The release note mentions a HIP tweak for CDNA GPUs. The binaries are available for download from the GitHub release page.
Why it matters
Developers can now run llama.cpp locally on more devices, including Macs with Apple Silicon and Ubuntu machines with GPU acceleration.