TY - RPRT TI - Soft Prompt Threats: Attacking Safety Alignment and Unlearning in Open-Source LLMs through the Embedding Space AU - Leo Schwinn AU - David Dobre AU - Sophie Xhonneux AU - Gauthier Gidel AU - Stephan Gunnemann PY - 2025 UR - https://arxiv.org/abs/2402.09063 ID - 2402.09063 ER -