Oossa

llama.cpp vb11391 adds CUDA tweak and new pre‑built binaries

The open‑source LLM runner releases version b11391 with a small CUDA code change and fresh binaries for macOS, iOS and several Ubuntu platforms.

NoteOossaPublished by Oossa: 1 min read

The ggml‑org team released llama.cpp version b11391 on 4 October 2026. The update moves a CUDA variable called blocks_per_col to where it is used, a change signed off by Adrien Gallouët. The release also bundles new pre‑compiled binaries for macOS Apple Silicon, macOS Intel, iOS, and multiple Ubuntu builds (CPU, Vulkan and CUDA 12). All files are available for download from the GitHub release page.

Why it matters

Developers can now run llama.cpp on more hardware out‑of‑the‑box, especially those using CUDA‑enabled GPUs.

Was this article useful?
Share

Read next

Oossa · Newsletter

The week in AI, explained

Every Monday: the stories worth knowing, in plain language. Free, no spam.