The llama.cpp project released build b10231, introducing support for the DSpark speculative decoding sidecar. This update allows dspark files to resolve similarly to other speculative sidecars, with the -hfd tag applying to them and auto-selection prioritizing dspark over dflash due to its extra Markov head.

The release provides binaries for macOS (Apple Silicon and Intel), iOS, Linux (Ubuntu x64, arm64, s390x with CPU, Vulkan, ROCm 7.2, OpenVINO, and SYCL backends), Android (arm64), Windows (CPU, OpenCL Adreno, CUDA 12/13, Vulkan, OpenVINO, SYCL, HIP), and openEuler (x86 and aarch64 with ACL Graph).

This update expands the available hardware acceleration options for llama.cpp users across multiple operating systems and GPU architectures.