# llama.cpp adds AVX512‑FP16 dot product support in b11262 release

> The open‑source llama.cpp library now uses AVX512‑FP16 to compute half‑precision dot products in full precision, improving CPU performance.

Oossa · 2026-09-29 · https://oossa.com/en/llama-cpp-adds-avx512-fp16-dot-product-support-in-b11262-release

The ggml‑org team released llama.cpp version b11262 on Sep 29, 2026. The update adds AVX512‑FP16 support, which does half‑precision (f16) dot products but accumulates them in single‑precision (f32) for better accuracy. It is a low‑level change that speeds up inference on CPUs that have the AVX512‑FP16 instruction set. The release also bundles new binaries for macOS, iOS, and several Linux builds.

## The facts

- Version b11262 released Sep 29, 2026
- Adds AVX512‑FP16 dot‑product accumulation in f32

## Why it matters

It lets developers run LLMs faster on modern CPUs without losing numerical quality.

## Sources & references

1. [ggml-org/llama.cpp b11262](https://github.com/ggml-org/llama.cpp/releases/tag/b11262) – llama.cpp, 2026-09-29

Last updated: 2026-09-29
