The llama.cpp project has released build b10259, which introduces the ability to allow the reshape of tensors during the loading process.
This update includes binaries for macOS (Apple Silicon and Intel), Linux (Ubuntu with CPU, Vulkan, ROCm, OpenVINO, and SYCL backends), Android, Windows (CPU, CUDA 12/13, Vulkan, OpenCL, HIP, OpenVINO, and SYCL), and openEuler.
The release also provides the standalone UI package for users.