The llama.cpp project has released build b10688, which includes a pull request adding fast vector (fa-vec) tunings for the M2 architecture.

This update provides binaries and frameworks across multiple platforms and hardware configurations. macOS builds are available for Apple Silicon and Intel, while iOS users can access an XCFramework. Linux support covers CPU, Vulkan, ROCm 7.14, OpenVINO, and SYCL variants for x64, arm64, and s390x architectures. Windows releases include CPU, OpenCL Adreno, Vulkan, ROCm 7.14, OpenVINO, SYCL, and CUDA 12/13 support for both x64 and arm64. Android arm64 binaries are also provided.

The release enables users to run llama.cpp on a wide range of devices and accelerators using the latest build artifacts.