Oossa

llama.cpp v0.1.15 adds CUDA out‑of‑bounds fix and new binaries

The ggml‑org release on Oct 8, 2026 patches a CUDA memory issue and provides pre‑built binaries for macOS, iOS and several Linux platforms.

NoteBy Published by Oossa: 1 min read

The open‑source llama.cpp library released version b11511 on Oct 8, 2026. It fixes a CUDA out‑of‑bounds read bug in the MMQ implementation. The update also adds ready‑to‑run binaries for macOS (Apple Silicon and Intel), iOS, and Ubuntu on x64, arm64 and s390x, plus Vulkan and CUDA‑12.8 builds.

Why it matters

Developers can now run llama.cpp on more hardware without crashes, especially those using NVIDIA GPUs.

Was this article useful?
Share

Read next

Oossa · Newsletter

The week in AI, explained

Every Monday: the stories worth knowing, in plain language. Free, no spam.