# llama.cpp update fixes sampling graph instability

> The Oct 9 2026 release makes the backend sampling graph static across ubatches, improving reliability on all supported platforms.

Oossa · 2026-10-09 · https://oossa.com/en/llama-cpp-update-fixes-sampling-graph-instability

The ggml‑org/llama.cpp project released version b11530 on 2026‑10‑09. The change keeps the backend sampling graph static across ubatches, so the graph no longer changes shape during reserve and decode steps. This fixes crashes that happened when the graph topology shifted after a reserve build. The fix applies to all listed builds – macOS, Linux, Android, Windows and openEuler – and to all hardware backends such as CUDA, Vulkan and OpenCL.

## The facts

- Release tag b11530 published on Fri Oct 09 2026
- Fix makes sampling graph static across ubatches, preventing graph‑topology changes

## Why it matters

Developers using llama.cpp will see fewer runtime errors when running multi‑output generations on any supported CPU or GPU.

## Sources & references

1. [ggml-org/llama.cpp b11530](https://github.com/ggml-org/llama.cpp/releases/tag/b11530) – llama.cpp, 2026-10-09

Last updated: 2026-10-09
