The llama.cpp project has released version b11147, which introduces a new binary kernel for OpenCL. This update specifically adds support for the A8 Q6_K non-MoE dp4a configuration.

  • Adds an A8 Q6_K non-MoE dp4a binary kernel for OpenCL.
  • Provides binaries for macOS (Apple Silicon and Intel), iOS, Linux (CPU, Vulkan, CUDA, ROCm, OpenVINO, SYCL, Snapdragon), Android, Windows (CPU, Adreno, CUDA, Vulkan, OpenVINO, SYCL, ROCm), and openEuler.

This release enables users to run llama.cpp on a wide range of hardware platforms with the latest kernel optimizations.