TY - RPRT TI - SAFE: Stable Alignment Finetuning with Entropy-Aware Predictive Control for Reinforcement Learning from Human Feedback (RLHF) AU - Dipan Maity PY - 2026 UR - https://arxiv.org/abs/2602.04651 ID - 2602.04651 ER -