# llama.cpp v b11531 adds new binaries and API refactor

> The open‑source LLM runtime releases macOS, iOS and Linux builds, plus a chat API overhaul.

Oossa · 2026-10-09 · https://oossa.com/en/llama-cpp-v-b11531-adds-new-binaries-and-api-refactor

The ggml‑org team published llama.cpp release b11531. It refactors the chat API and bundles pre‑built binaries for macOS (Apple Silicon and Intel), iOS, and several Ubuntu variants including CPU, Vulkan and CUDA 12.8. The download links are on the GitHub release page. Users can now run LLMs locally on more platforms without compiling themselves.

## The facts

- Release tag b11531 published on 2026-10-09
- Binaries provided for macOS arm64, macOS x64, iOS XCFramework, Ubuntu x64/arm64/s390x (CPU, Vulkan, CUDA 12.8)

## Why it matters

Developers can run Llama models locally on more devices today, avoiding build steps and supporting GPU acceleration on Ubuntu.

## Sources & references

1. [ggml-org/llama.cpp b11531](https://github.com/ggml-org/llama.cpp/releases/tag/b11531) – llama.cpp, 2026-10-09

Last updated: 2026-10-09
