The llama.cpp project has released version b11260, introducing support for FP32 GELU_ERF and GEGLU_ERF activation functions in the Hexagon backend. This update also includes optimizations to reduce register pressure within the associated kernels.

  • Added FP32 GELU_ERF and GEGLU_ERF support for Hexagon.
  • Reduced register pressure in hexagon kernels.
  • macOS Apple Silicon, Intel, and iOS binaries are available.
  • Linux builds cover CPU, Vulkan, CUDA 12/13, ROCm 10.0, OpenVINO, SYCL, and Snapdragon platforms.
  • Android arm64 builds include support for Snapdragon Hexagon NPU.
  • Windows releases provide CPU, OpenCL Adreno, CUDA 12/13, Vulkan, OpenVINO, SYCL, and ROCm 10.0 options.

This release expands hardware compatibility for specific activation functions and improves efficiency on supported accelerators.