arXiv · 2606.16794
LLM-Based Visual Explanation Evaluation Framework for Assessing the Explainability of Facial Skin Disease Classification Models
Abstract
This study proposes a domain-specific LLM-based Visual Explanation Evaluation Framework for assessing visual attention explanations in facial skin disease diagnosis. While previous studies have primarily focused on improving classification performance, relatively few studies have systematically examined whether visual explanations are grounded in clinically relevant lesion regions. In this study, an image-driven visual attention generation algorithm was developed to produce lesion-focused attention maps from facial skin disease images. Unlike conventional gradient-based explainability methods, the proposed approach combines color saliency, facial spatial priors, Gaussian smoothing, and attention overlay visualization to generate clinically interpretable attention maps. Furthermore, an LLM-as-a-Judge evaluation framework was designed using GPT-5.5, Gemini 3.5 Flash, and Claude Sonnet 4.6 to assess the generated visual explanations from the perspectives of lesion localization and explanation trustworthiness.
Explore related subjects
Keep this discovery
Gyuyeon Na. 2026-06-15. LLM-Based Visual Explanation Evaluation Framework for Assessing the Explainability of Facial Skin Disease Classification Models. https://arxiv.org/abs/2606.16794
Cite the original work for its findings. Save a collection to share your selection of sources.