arXiv · 2508.19288
Tricking LLM-Based NPCs into Spilling Secrets
Abstract
Large Language Models (LLMs) are increasingly used to generate dynamic dialogue for game NPCs. However, their integration raises new security concerns. In this study, we examine whether adversarial prompt injection can cause LLM-based NPCs to reveal hidden background secrets that are meant to remain undisclosed.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Kyohei Shiomi, Zhuotao Lian, Toru Nakanishi, Teruaki Kitasuka. 2025-08-25. Tricking LLM-Based NPCs into Spilling Secrets. https://arxiv.org/abs/2508.19288
Cite the original work for its findings. Save a collection to share your selection of sources.