Psychometric Validation
Treating AI person-perception as a measurement problem: construct validity, convergent and discriminant evidence, and criterion ceilings.
- VIDEO
Calibration as Co-Equal Criterion
A spoken walkthrough of why confidence–accuracy coupling belongs beside accuracy as a headline criterion, never composited into a single number.
- VIDEO
Sycophancy as Construct Invalidity
A spoken walkthrough of why sycophancy is a measurement pathology rather than a tone problem: an instrument that drifts toward what the subject wants to hear is not measuring the subject.
- VIDEO
The AI as Candidate Instrument
A spoken walkthrough of what it means to treat an AI system that judges people as a candidate psychometric instrument, and the validity battery it has to survive.
- ARTICLE
The AI as Candidate Instrument: A Full Validity Battery for Machine Judgments of Persons
If an AI system judges the people it talks to, it is a candidate instrument. This article specifies the five-component validity battery — convergence, discrimination, dose–response, calibration, and profile accuracy — needed to test one.
- ARTICLE
Calibration as Co-Equal Criterion: Confidence–Accuracy Coupling in Machine Judgments of Humans
A system that is confidently wrong about people is a different kind of object than one that is uncertainly wrong. This article argues calibration belongs beside accuracy as a headline criterion, never composited into one number.
- VIDEO
The Unmeasured Instrument
A spoken walkthrough of the argument: AI person-perception is a psychometric problem, and no benchmark yet validates whether machine judgments of people are accurate.
- ARTICLE
Sycophancy as Construct Invalidity: When the Instrument Measures Its Own Incentives
Sycophancy is not a tone problem but a measurement pathology: an instrument whose readings bend toward what the subject wants to hear is no longer measuring the subject.
- ARTICLE
The Unmeasured Instrument: AI Person-Perception as Psychometrics' Missing Validation Problem
AI systems now judge personality at scale, yet no benchmark validates whether those judgments are accurate. This piece argues that person-perception is a psychometric problem — and sketches what an independent validation standard would require.