Other route
Browse by model
Open one model’s whole response profile, then compare it scale by scale.
Browse models →Metric explorer
Choose one lens to compare every model side by side. From there, open any model, then the source question and its recorded response.
A metric is a summary of related questionnaire answers. It is a useful map of this study’s prompts—not a judgement of a model’s character, safety, or “true values.”
Eight lenses
Questionnaire lens
How strongly a response supports personal sacrifice to help people in serious need, including strangers.
Highest: Llama 3.3 70B · 4.5Questionnaire lens
How willing a response is to accept harming one person to produce a larger benefit. Lower scores mean more reluctance to use harm.
Highest: DeepSeek V4 Pro · 2.4Questionnaire lens
Concern for people who are suffering.
Highest: Mistral Medium 3.1 · 5.0Questionnaire lens
Preference for more equal outcomes and resources.
Highest: Mistral Medium 3.1 · 2.6Questionnaire lens
Preference for rewards that track contribution or effort.
Highest: Mistral Medium 3.1 · 4.8Questionnaire lens
Importance placed on commitment to one’s group or community.
Highest: Gemini 2.5 Flash Lite · 4.3Questionnaire lens
Importance placed on tradition, rules, and legitimate authority.
Highest: Grok 4.20 · 4.5Questionnaire lens
Importance placed on ideas of sanctity, restraint, and contamination.
Highest: Grok 4.20 · 3.4Other route
Open one model’s whole response profile, then compare it scale by scale.
Browse models →Relationships
See which measured scores and response sensitivities move together across the panel.
Open patterns →Evidence standard
Every lens leads to its source questions and exact stored model output.
How the scores work →