TY - RPRT TI - Reducing the Probability of Undesirable Outputs in Language Models Using Probabilistic Inference AU - Stephen Zhao AU - Aidan Li AU - Rob Brekelmans AU - Roger Grosse PY - 2025 UR - https://arxiv.org/abs/2510.21184 ID - 2510.21184 ER -