The ggml‑org team published llama.cpp release b11277. It patches a buffer‑reset issue that caused memory leaks. The update also ships pre‑built binaries for macOS (Apple Silicon and Intel), iOS, and several Ubuntu flavors, including Vulkan and CUDA GPU options. Users can download the files directly from the release page.
Why it matters
Fixing the leak improves stability for developers running LLMs locally, and the ready‑made binaries make it easier to try llama.cpp on more platforms.