Oossa

llama.cpp update fixes CUDA version check regression

The Oct 8, 2026 release corrects a version‑guard bug that slowed CUDA arg‑sort on newer CCCL releases.

NoteBy Published by Oossa: 1 min read

The llama.cpp project released version b11509 on Oct 8, 2026. It fixes a CUDA guard that mis‑identified CCCL version 4.x as older, causing a silent performance drop in the argsort operation. The fix replaces the component‑wise check with a single packed integer comparison, restoring the fast strided‑iterator path. The change was tested on an RTX 4070 with CUDA 13.4 and passed all backend tests.

Why it matters

Developers using llama.cpp on modern CUDA setups will see faster sorting performance without changing their code.

Was this article useful?
Share

Read next

Oossa · Newsletter

The week in AI, explained

Every Monday: the stories worth knowing, in plain language. Free, no spam.