Recognition, Simulation, and Refusal: A Contamination-Aware Study of Classic Psychological Effects in LLM Agents
PsyAgentBench separates LLM mimicry from genuine psychological bias using a factorial design.
Researchers introduced PsyAgentBench to distinguish between an LLM's ability to simulate human psychological response patterns and its actual possession of those biases. By testing models with both labeled and blind prompts, as well as canonical and counterfactual scenarios, the study aims to isolate lexical contamination from true behavioral traits.