# llama.cpp version b11496 adds SYCL bulk upload support

> The open‑source llama.cpp library released version b11496 on Oct 8 2026, introducing a SYCL‑based bulk model loading path via a pinned ring buffer.

Oossa · 2026-10-08 · https://oossa.com/en/llama-cpp-version-b11496-adds-sycl-bulk-upload-support

The ggml‑org team published llama.cpp b11496 on Oct 8 2026. The update adds a SYCL stage that moves bulk model uploads through a pinned ring buffer, speeding up loading on compatible GPUs. Binaries for macOS (Apple Silicon and Intel), iOS, and various Ubuntu builds (CPU, Vulkan, CUDA 12) are available for download.

## The facts

- Release tag: b11496 (Oct 8 2026)
- New SYCL bulk upload via pinned ring buffer

## Why it matters

Developers can load large language models faster on SYCL‑compatible hardware, reducing startup time for apps that use llama.cpp.

## Sources & references

1. [ggml-org/llama.cpp b11496](https://github.com/ggml-org/llama.cpp/releases/tag/b11496) – llama.cpp, 2026-10-08

Last updated: 2026-10-08
