TY - RPRT TI - Challenges in Ensuring AI Safety in DeepSeek-R1 Models: The Shortcomings of Reinforcement Learning Strategies AU - Manojkumar Parmar AU - Yuvaraj Govindarajulu PY - 2025 UR - https://arxiv.org/abs/2501.17030 ID - 2501.17030 ER -