ggml‑org has published the b11539 release of llama.cpp, the lightweight LLM runner that works on many devices. New binary archives are available for macOS Apple Silicon, macOS Intel, iOS XCFramework, and several Ubuntu builds covering CPU, Vulkan and CUDA 12.8. The download links are posted on the GitHub release page.
Why it matters
Developers can now run llama.cpp locally on more hardware without building it themselves.