Oossa

llama.cpp release b11498 adds CUDA kernel fix and new binaries

The open‑source LLM runtime updates its CUDA code and provides fresh pre‑built binaries for macOS, iOS and several Linux platforms.

NoteBy Published by Oossa: 1 min read

The ggml‑org team released llama.cpp version b11498 on 2026‑10‑08. The update fixes a CUDA kernel issue that occurred when certain tensor dimensions exceeded GPU grid limits. It also bundles new pre‑compiled binaries for Apple Silicon, Intel macOS, iOS, and multiple Ubuntu configurations, including CPU, Vulkan and CUDA 12.8 builds.

Why it matters

Developers can now run llama.cpp on more GPU setups without crashes and get ready‑to‑run binaries for their platform.

Was this article useful?
Share

Read next

Oossa · Newsletter

The week in AI, explained

Every Monday: the stories worth knowing, in plain language. Free, no spam.