TY - RPRT TI - Evaluating the performance and fragility of large language models on the self-assessment for neurological surgeons AU - Krithik Vishwanath AU - Anton Alyakin AU - Mrigayu Ghosh AU - Jin Vivian Lee AU - Daniel Alexander Alber AU - Karl L. Sangwon AU - Douglas Kondziolka AU - Eric Karl Oermann PY - 2025 UR - https://arxiv.org/abs/2505.23477 ID - 2505.23477 ER -