ggml-org released llama.cpp version b11537. The update reorders graph get_rows for embeddings and fixes Gemma4 embedding handling. It also adds a TODO note for LoRA support and corrects raw embedding paths. Pre‑compiled binaries are now available for macOS Apple Silicon, macOS Intel, iOS, and various Ubuntu builds, including Vulkan‑enabled versions.
Why it matters
Developers can now run llama.cpp locally on more platforms with corrected embedding support, reducing setup hassle.