The llama.cpp project has released build b11067, which introduces a fused GDN + copy operation for the WebGPU backend.
- The release includes binaries for macOS (Apple Silicon and Intel), Linux (Ubuntu with CPU, Vulkan, CUDA 12/13, ROCm 10.0, OpenVINO, SYCL), Windows (CPU, OpenCL Adreno, CUDA 12/13, Vulkan, OpenVINO, SYCL, ROCm 10.0), Android, and iOS.
- KleidiAI support for macOS Apple Silicon is disabled in this build.
- openEuler support is also disabled.
This update provides users with updated binaries across multiple platforms and accelerators, alongside the specific WebGPU optimization.