ANES 2024: Comparing Synthetic and Later Human Answers
An ANES pre/post-election comparison tests 300 matched people on 28 later answers, with separate results for response format and participant grounding.
Experiments, benchmarks, and technical papers from Minds Research Lab on synthetic respondents, persona-conditioned AI, evaluation methods, and the limits of simulated evidence.
An ANES pre/post-election comparison tests 300 matched people on 28 later answers, with separate results for response format and participant grounding.
A CES comparison evaluates 300 matched respondents on 27 binary, ordinal and multiselect questions, separating aggregate fit from individual prediction.
A target-separated GSS pilot compares four response conditions across 360 people and 17 questions, separating probability elicitation from profile value.
A five-country PISA comparison evaluates 600 students and ten held-out questions, with weighted distribution errors and explicit country and language limits.
What five public survey datasets reveal about synthetic audience accuracy, including a 301-Mind Gen Z production test, and where profile grounding helps.
An outcome-blind validation compared 903 responses from 301 persistent Gen Z Minds with published UK Food Standards Agency survey distributions. Aggregate approximation reached 93.99%.