# llama.cpp update adds OpenCL fixes for Q4_K weight limits

> The llama.cpp library release b11558 (Oct 11, 2026) improves OpenCL handling of Q4_K weight images and fixes several kernel bugs.

Oossa · 2026-10-11 · https://oossa.com/en/llama-cpp-update-adds-opencl-fixes-for-q4-k-weight-limits

The open‑source llama.cpp project released version b11558 on Oct 11, 2026. The update fixes a bug that caused OpenCL kernels to exceed device image limits when using Q4_K quantized weights. It also adds a fallback path for devices that hit those limits, and corrects broadcasting and RMS‑norm calculations for Q4_0 kernels. The changes were co‑authored by Hongqiang Wang of Qualcomm.

## The facts

- Release version b11558 published on 2026-10-11
- Fixes OpenCL image limit for Q4_K dense bin kernels

## Why it matters

Developers using llama.cpp on GPUs can now run Q4_K quantized models on more hardware without crashes.

## Sources & references

1. [ggml-org/llama.cpp b11558](https://github.com/ggml-org/llama.cpp/releases/tag/b11558) – llama.cpp, 2026-10-11

Last updated: 2026-10-11
