arXiv · 2602.06714
PrefIx: Understand and Adapt to User Preference in Human-Agent Interaction
Abstract
LLM-based agents can complete tasks correctly yet still frustrate users through poor interaction patterns, such as excessive confirmations, opaque reasoning, or misaligned pacing. Current benchmarks evaluate task accuracy but overlook how agents interact: whether they infer preferences from implicit cues, adapt dynamically, or maintain fine-grained interaction quality. We introduce Prefix, a configurable environment that evaluates both what agents accomplish and how they interact. Central to Prefix is the Interaction-as-a-Tool (IaaT) paradigm, which treats interaction behaviors as structured tool calls, unifying them with existing evaluation frameworks. We define 31 preference settings across 14 attributes and formalize user experience (UX) as a core metric alongside task accuracy. A composite LLM-as-a-Judge mechanism across seven UX dimensions achieves strong aggregate reliability (ICC > 0.79), high internal consistency (alpha = 0.943), and human correlation (rho = 0.52-0.78). Preference-aware agents show 7.6% average UX improvement and 18.5% gain in preference alignment. Our work is openly accessible.
Explore related subjects
Keep this discovery
Jialin Li, Zhenhao Chen, Hanjun Luo, Hanan Salam. 2026-02-06. PrefIx: Understand and Adapt to User Preference in Human-Agent Interaction. https://arxiv.org/abs/2602.06714
Cite the original work for its findings. Save a collection to share your selection of sources.