Ling-3.0-flash-VL is a new model built on the Ling-3.0-flash architecture that introduces visual understanding and visual agent capabilities.

The model performs well across several domains, including visual perception, STEM reasoning, document intelligence, multimodal agent tasks, frontend coding, and medical report interpretation.

This release expands the base model's utility by enabling it to process and interact with visual data effectively.