Aloe-Vision: Robust Vision-Language Models for Healthcare
This work introduces Aloe-Vision, a family of open-source large vision-language models (7B and 72B) trained on the newly released Aloe-Vision-Data dataset to address data scarcity and robustness issues in healthcare AI. The authors demonstrate that their high-quality training mixture yields significant performance gains over baselines while maintaining general capabilities.