Oossa

llama.cpp release adds MUSA support and updates CUDA versions

The open‑source Llama.cpp library now handles MUSA GPUs and includes builds for CUDA 12.8 and 13.4.

NoteOossaPublished by Oossa: 1 min read

The ggml‑org team released a new version of llama.cpp on September 30, 2026. It fixes a problem where the MUSA GPU driver did not define CUDA_ARCH, causing some kernels to compile empty. The patch adds proper architecture detection and removes redundant checks. The release also lists many pre‑built binaries, now covering CUDA 12.8 and 13.4 on Linux and Windows, plus support for various ARM and Apple Silicon targets.

Why it matters

Developers can now run Llama.cpp on MUSA GPUs and newer CUDA versions without manual patches.

Was this article useful?
Share

Read next

Oossa · Newsletter

The week in AI, explained

Every Monday: the stories worth knowing, in plain language. Free, no spam.