The llama.cpp project has released version b11045, which introduces support for the ROLL operation within the Hexagon backend.

  • The release includes binaries for macOS (Apple Silicon and Intel), Linux (Ubuntu with CPU, Vulkan, CUDA 12/13, ROCm 10.0, OpenVINO, and SYCL), Windows (CPU, OpenCL Adreno, CUDA 12/13, Vulkan, OpenVINO, SYCL, and ROCm 10.0), Android arm64, and openEuler.
  • An iOS XCFramework and a standalone UI build are also provided.
  • KleidiAI support for macOS Apple Silicon is currently disabled in this release.