# llama.cpp v0.1.0 release adds GPU subtraction tweak and new binaries

> The open‑source llama.cpp library released version b11270 with a GPU subtraction optimization and pre‑built binaries for macOS, iOS and Ubuntu.

Oossa · 2026-09-30 · https://oossa.com/en/llama-cpp-v0-1-0-release-adds-gpu-subtraction-tweak-and-new-binaries

The ggml‑org team posted llama.cpp version b11270 on Sep 30, 2026. It tweaks a GPU operation used by AMD’s HIP platform, replacing a slower byte‑subtraction instruction with a faster one. The release also bundles fresh binary packages for macOS (Apple Silicon and Intel), iOS, and several Ubuntu builds, including Vulkan‑enabled GPU versions.

## The facts

- Release date: Wed Sep 30 2026
- Adds HIP GPU instruction change __vsubss4 → __vsub4

## Why it matters

The GPU tweak can speed up inference on AMD hardware, and the ready‑to‑run binaries make it easier for developers to try llama.cpp on more platforms.

## Sources & references

1. [ggml-org/llama.cpp b11270](https://github.com/ggml-org/llama.cpp/releases/tag/b11270) – llama.cpp, 2026-09-30

Last updated: 2026-09-30
