TY - RPRT TI - Interpretable Steering of Large Language Models with Feature Guided Activation Additions AU - Samuel Soo AU - Chen Guang AU - Wesley Teng AU - Chandrasekaran Balaganesh AU - Tan Guoxian AU - Yan Ming PY - 2025 UR - https://arxiv.org/abs/2501.09929 ID - 2501.09929 ER -