arXiv · 2505.18792
On the Dual-Use Dilemma in Physical Reasoning and Force
Abstract
Humans learn how and when to apply forces in the world via a complex physiological and psychological learning process. Attempting to replicate this in vision-language models (VLMs) presents two challenges: VLMs can produce harmful behavior, which is particularly dangerous for VLM-controlled robots which interact with the world, but imposing behavioral safeguards can limit their functional and ethical extents. We conduct two case studies on safeguarding VLMs which generate forceful robotic motion, finding that safeguards reduce both harmful and helpful behavior involving contact-rich manipulation of human body parts. Then, we discuss the key implication of this result--that value alignment may impede desirable robot capabilities--for model evaluation and robot learning.
Explore related subjects
Keep this discovery
William Xie, Enora Rice, Nikolaus Correll. 2025-05-24. On the Dual-Use Dilemma in Physical Reasoning and Force. https://arxiv.org/abs/2505.18792
Cite the original work for its findings. Save a collection to share your selection of sources.