OossaAI is evolving fast. We explain it simply.

llama.cpp release b11268 adds OpenCL fix for Adreno GPUs

The open‑source LLM runner updates its OpenCL code to correct tensor handling on Adreno GPUs and ships new binaries for macOS, iOS and Linux.

NoteOossa1 min read

The ggml‑org team released llama.cpp version b11268. The update fixes a bug in the OpenCL driver that handled tensors for the q5_K Adreno GPU kernel. New pre‑built binaries are provided for macOS Apple Silicon, Intel Macs, iOS, and several Linux configurations including CPU, Vulkan and CUDA 12.8.

Why it matters

The fix improves performance and stability for developers running LLMs on Android devices with Adreno GPUs.

Was this article useful?
Share

Read next

Oossa · Newsletter

The week in AI, explained

Every Monday: the stories worth knowing, in plain language. Free, no spam.

Sources & references

#SourceOutletDateKey takeaway
1ggml-org/llama.cpp b11268 ↗llama.cppSep 30, 2026<details open> opencl: fix get_tensor for q5_K adreno gemm_nonshuffle kernel (#29555) </details> **Website:** - <https://llama.app> **Attest

1 sources

Last updated: ·Markdown·llms.txt