Oossa

llama.cpp update adds host tensor view support and re‑enables K2 Horizon tensors

The Oct 9 2026 release fixes a crash when using KV‑cache splits and restores support for K2 Horizon hardware.

NoteBy Published by Oossa: 1 min read

The ggml‑org team released llama.cpp version b11528 on 2026‑10‑09. The update lets the library treat views of host‑allocated tensors as normal nodes, stopping a previous assert error that appeared with the –split-mode tensor option. It also brings back SM tensor support for the K2 Horizon accelerator. The change is included in the official GitHub release and applies to all listed platforms, from macOS to Windows and Linux.

Why it matters

Developers can now run llama.cpp with split‑mode KV caching on more hardware without crashes, expanding usable configurations for local LLM inference.

Was this article useful?
Share

Read next

Oossa · Newsletter

The week in AI, explained

Every Monday: the stories worth knowing, in plain language. Free, no spam.