TY - RPRT TI - Faithful-Patchscopes: Understanding and Mitigating Model Bias in Hidden Representations Explanation of Large Language Models AU - Xilin Gong AU - Shu Yang AU - Zehua Cao AU - Lynne Billard AU - Di Wang PY - 2026 UR - https://arxiv.org/abs/2602.00300 ID - 2602.00300 ER -