arXiv · 2608.05166
Conditional Cognitive Biases in LLMs: How Biased User Turns Modulate In-Context Reasoning
Abstract
We present an evaluation of cognitive bias expression in state-of-the-art instruction-tuned LLMs under realistic multi-turn interaction settings. Our work introduces a novel three-condition experimental framework that disentangles the effect of exposure to a biased user turn from the effect of the turn's semantic content, alongside a benchmark of 24,300 jury-validated user prompts spanning all 81 cells of a 9x9 target-human bias interaction matrix. Across eight frontier LLMs, we find that biased conversational context systematically increases bias expression relative to zero-shot baselines in 6 of 8 models. We identify two competing behavioral dynamics underlying this effect: conversational exposure to biased reasoning generally amplifies downstream bias tendencies, while explicitly stated bias cues often trigger alignment-related suppression behaviors that reduce overt bias expression. We release our framework, codebase, and dataset to support future research on context-conditioned cognitive biases and behavioral adaptation in LLMs.
Explore related subjects
Keep this discovery
Sachini Weerasekara, Sagar Kamarthi, Jacqueline Isaacs. 2026-05-26. Conditional Cognitive Biases in LLMs: How Biased User Turns Modulate In-Context Reasoning. https://arxiv.org/abs/2608.05166
Cite the original work for its findings. Save a collection to share your selection of sources.