---
title: "GSS 2024: Preserving Uncertainty Improves Survey Fit | Minds"
canonical_url: "https://getminds.ai/research/gss-2024-synthetic-survey-validation"
last_updated: "2026-09-09T03:02:31.583Z"
meta:
  description: "A target-separated GSS pilot compares four response conditions across 360 people and 17 questions, separating probability elicitation from profile value."
  "og:description": "A target-separated GSS pilot compares four response conditions across 360 people and 17 questions, separating probability elicitation from profile value."
  "og:title": "GSS 2024: Preserving Uncertainty Improves Survey Fit | Minds"
  "twitter:description": "A target-separated GSS pilot compares four response conditions across 360 people and 17 questions, separating probability elicitation from profile value."
  "twitter:title": "GSS 2024: Preserving Uncertainty Improves Survey Fit | Minds"
---

Minds

September 5, 2026·Validation·Minds Team # **GSS 2024: Preserving Uncertainty Improves Survey Fit** A target-separated GSS pilot compares four response conditions across 360 people and 17 questions, separating probability elicitation from profile value. The GSS 2024 pilot reached 92.47% aggregate approximation with structured-profile probability answers. Its mean distribution error was 7.53 percentage points across 17 questions and 360 matched respondents. The useful finding was not simply a high final score: collecting probabilities substantially improved distribution fit over asking for one forced answer. ## What was compared The evaluation used real GSS attitudes and behavior responses as held-out targets. Four conditions distinguished the answer format from the context supplied to the model: generic forced choice, generic probability, demographic probability, and structured-profile probability. Candidates were sealed before the evaluation targets were opened. The profile contained structured survey evidence, not the complete production Mind representation. The experiment ran in an isolated harness. Its result supports the tested response method, not identical accuracy for every production audience. ## Results across the four conditions_### **Mean item absolute error** Reported results for this study. The table and discussion retain the study scope, baselines and uncertainty._- Generic forced choice**25.74**