---
title: "PRISM Alignment: Testing Participant Fit in… | Minds"
canonical_url: "https://getminds.ai/research/prism-alignment-participant-fit-validation"
last_updated: "2026-09-09T04:00:58.022Z"
meta:
  description: "A 299-person, 869-item comparison measures whether grounded Minds better reflect participants' values and reasons, with a blinded judge and clear task limits."
  "og:description": "A 299-person, 869-item comparison measures whether grounded Minds better reflect participants' values and reasons, with a blinded judge and clear task limits."
  "og:title": "PRISM Alignment: Testing Participant Fit in… | Minds"
  "twitter:description": "A 299-person, 869-item comparison measures whether grounded Minds better reflect participants' values and reasons, with a blinded judge and clear task limits."
  "twitter:title": "PRISM Alignment: Testing Participant Fit in… | Minds"
---

Minds

September 5, 2026·Validation·Minds Team # **PRISM Alignment: Testing Participant Fit in Held-Out Answers** A 299-person, 869-item comparison measures whether grounded Minds better reflect participants' values and reasons, with a blinded judge and clear task limits. In a confirmatory comparison using the external PRISM Alignment Dataset, full Minds won 32.6% of blinded participant-fit evaluations, compared with 25.7% for demographic prompts and 20.5% for generic prompts. The strongest gains concerned values expression, specificity and lower genericness. This PRISM dataset is an external benchmark. It is distinct from the proprietary Minds PRISM source-modeling engine. ## Participants, targets and conditions The study evaluated 299 UK participants and 869 participant-domain items. Three conditions produced 2,607 scored outputs: full Minds through the production path, demographic-only prompts, and generic prompts. Selection was outcome blind and real participant answers were held out from generation. A single blinded LLM judge evaluated condition-anonymous, length-matched comparisons. This is automated participant-fit assessment, not independent human-rater validation. ## What the judge preferred_### **Share of evaluated items** Reported results for this study. The table and discussion retain the study scope, baselines and uncertainty._- Full Minds winner**32.5662%**