TY - RPRT TI - Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs AU - Keenan Pepper AU - Alex McKenzie AU - Florin Pop AU - Stijn Servaes AU - Martin Leitgab AU - Mike Vaiana AU - Judd Rosenblatt AU - Michael S. A. Graziano AU - Diogo de Lucena PY - 2026 UR - https://arxiv.org/abs/2602.10352 ID - 2602.10352 ER -