arXiv · 2603.11597
Performance Evaluation of Open-Source Large Language Models for Assisting Pathology Report Writing in Japanese
Abstract
The performance of large language models (LLMs) for supporting pathology report writing in Japanese remains unexplored. We evaluated seven open-source LLMs from three perspectives: (A) generation and information extraction of pathology diagnosis text following predefined formats, (B) correction of typographical errors in Japanese pathology reports, and (C) subjective evaluation of model-generated explanatory text by pathologists and clinicians. Thinking models and medical-specialized models showed advantages in structured reporting tasks that required reasoning and in typo correction. In contrast, preferences for explanatory outputs varied substantially across raters. Although the utility of LLMs differed by task, our findings suggest that open-source LLMs can be useful for assisting Japanese pathology report writing in limited but clinically relevant scenarios.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Masataka Kawai, Singo Sakashita, Shumpei Ishikawa, Shogo Watanabe, Anna Matsuoka, Mikio Sakurai, Yasuto Fujimoto, Yoshiyuki Takahara, Atsushi Ohara, Hirohiko Miyake, Genichiro Ishii. 2026-03-12. Performance Evaluation of Open-Source Large Language Models for Assisting Pathology Report Writing in Japanese. https://arxiv.org/abs/2603.11597
Cite the original work for its findings. Save a collection to share your selection of sources.