Oossa

llama.cpp release b11493 adds grouped MoE XMX GEMM support

The open‑source llama.cpp library released version b11493 on Oct 8 2026, adding a new SYCL‑based grouped MoE XMX GEMM operation and binaries for many platforms.

NoteBy Published by Oossa: 1 min read

The ggml‑org team pushed llama.cpp version b11493 on Oct 8 2026. The update adds a SYCL implementation of grouped Mixture‑of‑Experts (MoE) XMX GEMM, a matrix multiply used in large language models. The release also bundles pre‑built binaries for macOS, iOS, Ubuntu (CPU, Vulkan and CUDA).

Why it matters

Developers can now run MoE‑based models on SYCL‑compatible GPUs, expanding hardware options for open‑source LLM workloads.

Was this article useful?
Share

Read next

Oossa · Newsletter

The week in AI, explained

Every Monday: the stories worth knowing, in plain language. Free, no spam.