Google Research·· 2026-04-03
Evaluating alignment of behavioral dispositions in LLMs
Evaluating alignment of behavioral dispositions in LLMs
AI summary
Google Research proposed situational judgement tests (SJT) adapted from psychological questionnaires such as IRI and ERQ to evaluate alignment between LLM behaviours and human consensus in real user–assistant scenarios.
Selection record
Threshold 60Official, first-handFirst 33Second 38
Not admittedSum of both 71 < twice the threshold 120
- Source tier
- Official, first-hand; this tier's threshold is 60
- Pre-filter
- passed:评估LLM行为倾向对齐的研究
A model scores each item twice, independently, against one written standard, out of 100. An item is admitted only when the two scores add up to twice the threshold. The threshold is set per source tier.
Source: Google Research · research.google