The llama.cpp project has released build b10696, which includes a pull request adding fast vector (fa-vec) tuning specifically for the Apple M3 Pro chip.

This update addresses issue #27668 to improve performance on this specific hardware configuration. The release provides binaries for macOS, iOS, Linux, Windows, Android, and openEuler across various backends including CPU, CUDA, ROCm, Vulkan, OpenVINO, and SYCL.

The fa-vec tuning is intended to optimize inference speed or efficiency for users running llama.cpp on devices equipped with the M3 Pro processor.