TY - RPRT TI - Lyapunov-Guided Self-Alignment: Test-Time Adaptation for Offline Safe Reinforcement Learning AU - Seungyub Han AU - Hyungjin Kim AU - Jungwoo Lee PY - 2026 UR - https://arxiv.org/abs/2604.26516 ID - 2604.26516 ER -