arXiv · 2607.00403
A Penny for Your Prompts: Experiments Detecting and Mitigating LLM Usage by Survey Respondents
Abstract
Large language models are increasingly used by participants on crowdsourcing platforms when responding to surveys, potentially undermining the validity of collected data. Our study aims to quantify the prevalence of this behavior and investigate methods to detect and prevent it. In a series of surveys (N = 250), we examined conditions such as platform choice, survey length, requests not to use AI, and disabling copy-paste functionality. We were able to identify distinct characteristics of LLM-assisted responses and found that their frequency varied widely, from under 10% on Prolific to over 80% on Mechanical Turk. Mitigation measures reduced LLM usage but did not necessarily improve data quality. No participants employed browser-use agents at the time of our survey, but we report on our own detection experiments. We recommend that researchers actively screen survey responses for LLM usage by recording and analyzing keystroke data and crafting instructions and questions aimed at AI.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Zane Xu, Nathan Malkin. 2026-07-01. A Penny for Your Prompts: Experiments Detecting and Mitigating LLM Usage by Survey Respondents. https://arxiv.org/abs/2607.00403
Cite the original work for its findings. Save a collection to share your selection of sources.