The ggml‑org team released llama.cpp version b11481 on 2026‑10‑08. The update adds support for Cohere2 vision models and includes pre‑built binaries for Apple Silicon, Intel macOS, iOS and several Ubuntu configurations. The files are available for download from the GitHub release page.
Why it matters
Developers can now run image‑aware AI models locally on macOS, iOS and Linux without cloud services.