The llama.cpp project has released version b11055, which introduces support for the GEGLU_QUICK activation function on Hexagon DSPs.
- The release includes binaries for macOS (Apple Silicon and Intel), Linux (Ubuntu x64, arm64, s390x), Windows, and Android.
- GPU acceleration is supported via CUDA 12.8/13.3, ROCm 10.0, Vulkan, OpenVINO, and SYCL on Linux, as well as CUDA, OpenCL, Vulkan, OpenVINO, and SYCL on Windows.
- macOS Apple Silicon builds with KleidiAI are currently disabled.
- openEuler support is also listed, though some configurations are marked as disabled.
This update provides users with broader hardware compatibility and optimized inference capabilities for Hexagon-based devices.