Oossa

llama.cpp 0.1.14 adds tiled Q4_K and Q6_K support for Hexagon

The open‑source Llama.cpp library now supports tiled Q4_K and Q6_K formats on Hexagon CPUs, with new binaries for macOS, iOS and Linux.

NoteBy Published by Oossa: 1 min read

The ggml‑org team released llama.cpp version b11487. It adds tiled Q4_K and Q6_K GET_ROWS support for Hexagon processors and improves handling of Q4_K views. Pre‑built binaries are provided for Apple Silicon, Intel macOS, iOS, and several Ubuntu builds (CPU and Vulkan). The update is available for download from the GitHub release page.

Why it matters

Developers can now run more efficient Llama models on Hexagon‑based devices such as Qualcomm chips, improving performance on supported mobile and embedded hardware.

Was this article useful?
Share

Read next

Oossa · Newsletter

The week in AI, explained

Every Monday: the stories worth knowing, in plain language. Free, no spam.