TY - RPRT TI - Offline Reinforcement Learning using Human-Aligned Reward Labeling for Autonomous Emergency Braking in Occluded Pedestrian Crossing AU - Vinal Asodia AU - Barkin Dagda AU - Yinglong He AU - Zhenhua Feng AU - Saber Fallah PY - 2026 UR - https://arxiv.org/abs/2504.08704 ID - 2504.08704 ER -