Logo image
Do participant-matched LLM personas approximate human survey data?
Journal article   Open access   Peer reviewed

Do participant-matched LLM personas approximate human survey data?

Ana Stojanov
Personality and individual differences, Vol.261, 113915
10/2026
Handle:
https://hdl.handle.net/10523/51424

Abstract

Artificial personas ChatGPT Individual differences Large language models Profile similarity Psychological measurement Survey simulation
Large language models can be asked to complete a survey as if they were a person. However, it is unclear how well such model implied responses resemble human data, especially when personas are conditioned on participant information. Human participants completed a questionnaire measuring 91 psychological constructs. For each participant (n = 177) we generated artificial personas using GPT-4o under two conditions: demographics-only versus demographics plus Big Five scores. We assessed correspondence between human and model implied responses across several metrics: construct means and dispersions, within-pair person-level distances, within-person profile similarity across constructs, and construct-wise human–persona correlations across participants. Adding Big Five information produced small improvements in mean alignment and individual-level similarity, large reductions in under-dispersion relative to humans, modest gains in construct-wise human–persona correspondence, and reduced bias, while RMSE and error variability improved mainly after averaging across multiple persona generations. Overall, model implied responses approximate some aggregate mean patterns across constructs but only modestly reproduce individual differences.
pdf
1-s2.0-S0191886926002795-main1.88 MBDownloadView
Published (Version of record) Open Access CC BY-NC-ND V4.0
url
https://doi.org/10.1016/j.paid.2026.113915View
Published (Version of record) Open CC BY-NC-ND V4.0

Metrics

1 Record Views

Details

Logo image