The llama.cpp project has released version b10638, which implements Vulkan support for the `cross_entropy_loss` and `cross_entropy_loss_back` functions.

  • The release includes pre-built binaries for macOS (Apple Silicon and Intel), Linux (Ubuntu x64, arm64, s390x), Windows (x64 and arm64), Android (arm64), and openEuler.
  • GPU acceleration is available via Vulkan, CUDA (12.4, 13.3, 13.4 preview), ROCm 7.14, OpenCL Adreno, SYCL, and OpenVINO across supported platforms.
  • macOS Apple Silicon builds with KleidiAI are currently disabled in this release.
  • An updated UI package is also provided alongside the core binaries.