The llama.cpp library released a patch on Sep 30, 2026, that blocks tensors whose size would overflow during padding. The change checks the raw size plus alignment before adding padding, rejecting unsafe tensors early. A new test case shows the fix catches a 2⁶⁴‑16‑byte tensor that previously slipped through. The update applies to all supported builds, from macOS to Windows and Linux, including CUDA and Vulkan versions.
Why it matters
Stopping this overflow prevents crashes or corrupted results when running very large models.