The llama.cpp project has released version b11260, introducing support for FP32 GELU_ERF and GEGLU_ERF activation functions in the Hexagon backend. This update also includes optimizations to reduce register pressure within the associated kernels.
- Added FP32 GELU_ERF and GEGLU_ERF support for Hexagon.
- Reduced register pressure in hexagon kernels.
- macOS Apple Silicon, Intel, and iOS binaries are available.
- Linux builds cover CPU, Vulkan, CUDA 12/13, ROCm 10.0, OpenVINO, SYCL, and Snapdragon platforms.
- Android arm64 builds include support for Snapdragon Hexagon NPU.
- Windows releases provide CPU, OpenCL Adreno, CUDA 12/13, Vulkan, OpenVINO, SYCL, and ROCm 10.0 options.
This release expands hardware compatibility for specific activation functions and improves efficiency on supported accelerators.