TY - RPRT TI - Lagrangian-based online safe reinforcement learning for state-constrained systems AU - Soutrik Bandyopadhyay AU - Shubhendu Bhasin PY - 2024 DO - 10.1016/j.automatica.2025.112458 UR - https://arxiv.org/abs/2305.12967 ID - 2305.12967 ER -