Simulating a population
Draft availableWhen personas are built from survey data, do their answers reproduce the population distribution, and can changes to that population be audited?
On 74 ANES demographic cells, calibrated personas had about one-third the vote-distribution error of naive repeated prompting and preserved within-cell variation. Direct readout, verbalized sampling, a calibrated label, and a training-data lookup matched or beat the persona arm on these coarse static cells. Steering was tested for ideology and religious attendance; it worked on DeepSeek and only partly on GPT-4o-mini.
Simpler methods are stronger baselines for static shares in well-surveyed groups. Persona value must be measured on individual variation and downstream interaction, with calibration and steering audits published alongside the result.
6/7, 7/7, 5/7 · latest assessment: justify publication