Oossa

llama.cpp release adds OpenCL support for Gemma‑4 GPU decoding

The Oct 11, 2026 update improves OpenCL performance and enables Gemma‑4 decoding on newer Qualcomm GPUs.

NoteBy Published by Oossa: 1 min read

The llama.cpp library released version b11559 on Oct 11, 2026. It adds OpenCL improvements for faster factor‑allocation and supports Gemma‑4 decoding on Qualcomm E4B and E2B GPUs. The update also optimizes DK64 and DK128 decode paths for various quantisation settings. The changes were co‑authored by Hongqiang Wang of Qualcomm.

Why it matters

Developers can now run Gemma‑4 models faster on compatible Qualcomm GPUs, expanding edge‑AI capabilities.

Was this article useful?
Share

Read next

Oossa · Newsletter

The week in AI, explained

Every Monday: the stories worth knowing, in plain language. Free, no spam.