TY - RPRT TI - Personality-Aware Reinforcement Learning for Persuasive Dialogue with LLM-Driven Simulation AU - Donghuo Zeng AU - Roberto Legaspi AU - Kazushi Ikeda PY - 2026 UR - https://arxiv.org/abs/2601.06877 ID - 2601.06877 ER -