The mtmd component in llama.cpp now supports the WebP image format through ffmpeg integration.
- Enables processing of WebP images within the mtmd framework.
- Relies on ffmpeg for decoding functionality.
The mtmd component in llama.cpp now supports the WebP image format through ffmpeg integration.
The llama.cpp project released build b10568, which updates the model handling to use the `ggml_rope_set_offset()` function. This change is partially applied to support DeepSeek 2.
The llama.cpp project has released version 0.2.0, which includes a synchronization with the ggml library bumped to version 0.21.0. This update introduces several backend improvements and new features across various hardware platforms.
The Celestium Engine has introduced several upgrades to improve SDXL performance on low-VRAM hardware. These changes focus on memory management, stability, and fallback mechanisms to ensure consistent generation even on constrained systems.
The llama.cpp project has released version b10549, which primarily introduces support for tensor splitting in LFM2 and LFM2MOE models.
The vLLM project released version 0.28.0rc2, which includes the DFlash2 speculative decoding feature.
We use cookies to measure traffic and improve the site. You can accept or decline analytics cookies. Privacy policy