The llama.cpp project has released version b11055, which introduces support for the GEGLU_QUICK activation function on Hexagon DSPs.

  • The release includes binaries for macOS (Apple Silicon and Intel), Linux (Ubuntu x64, arm64, s390x), Windows, and Android.
  • GPU acceleration is supported via CUDA 12.8/13.3, ROCm 10.0, Vulkan, OpenVINO, and SYCL on Linux, as well as CUDA, OpenCL, Vulkan, OpenVINO, and SYCL on Windows.
  • macOS Apple Silicon builds with KleidiAI are currently disabled.
  • openEuler support is also listed, though some configurations are marked as disabled.

This update provides users with broader hardware compatibility and optimized inference capabilities for Hexagon-based devices.