arXiv · 2509.18648
SPiDR: A Simple Approach for Zero-Shot Safety in Sim-to-Real Transfer
Abstract
Deploying reinforcement learning (RL) safely in the real world is challenging, as policies trained in simulators must face the inevitable sim-to-real gap. Robust safe RL techniques are provably safe, however difficult to scale, while domain randomization is more practical yet prone to unsafe behaviors. We address this gap by proposing SPiDR, short for Sim-to-real via Pessimistic Domain Randomization -- a scalable algorithm with provable guarantees for safe sim-to-real transfer. SPiDR uses domain randomization to incorporate the uncertainty about the sim-to-real gap into the safety constraints, making it versatile and highly compatible with existing training pipelines. Through extensive experiments on sim-to-sim benchmarks and two distinct real-world robotic platforms, we demonstrate that SPiDR effectively ensures safety despite the sim-to-real gap while maintaining strong performance.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Yarden As, Chengrui Qu, Benjamin Unger, Dongho Kang, Max van der Hart, Laixi Shi, Stelian Coros, Adam Wierman, Andreas Krause. 2025-09-23. SPiDR: A Simple Approach for Zero-Shot Safety in Sim-to-Real Transfer. https://arxiv.org/abs/2509.18648
Cite the original work for its findings. Save a collection to share your selection of sources.