TY - RPRT TI - Large Language Models are Vulnerable to Bait-and-Switch Attacks for Generating Harmful Content AU - Federico Bianchi AU - James Zou PY - 2024 UR - https://arxiv.org/abs/2402.13926 ID - 2402.13926 ER -