allenai/olmo-3.1-32b-instruct
Personality Summary
Big Five Personality
About this test
The most widely accepted personality framework in psychology. Five dimensions scored 1.0-5.0, with Neuroticism flipped to Emotional Stability so higher is always better. U.S. population averages: O 3.5, C 3.4, E 3.2, A 3.6, ES 3.1 (total 16.8 / 25.0).
- Openness โ Curiosity, imagination, preference for novelty.
- Conscientiousness โ Organization, discipline, reliability.
- Extraversion โ Sociability, assertiveness, positive energy.
- Agreeableness โ Cooperativeness, trust, empathy.
- Emotional Stability โ Calm, resilience (inverse of Neuroticism).
References: BFI-44 at Berkeley ยท John & Srivastava (1999)
Run-by-Run Breakdown
| Run | O | C | E | A | ES |
|---|---|---|---|---|---|
| 1 | 3.8 | 3.56 | 3.38 | 4 | 2.33 |
| 2 | 3.89 | 3.56 | 3.5 | 4.14 | 3 |
| 3 | 3.9 | 3.62 | 3.5 | 3.5 | 2.83 |
MBTI Dimensions
About this test
93 forced-choice questions across four dichotomies. In the U.S. population, the most common types are ISFJ (~14%) and ESFJ (~12%); the rarest are INFJ (~1.5%) and INTJ (~2%). About 73% of people lean Sensing over Intuition.
- E/I โ Extraversion vs. Introversion: where you direct energy.
- S/N โ Sensing vs. Intuition: concrete facts vs. patterns and possibilities.
- T/F โ Thinking vs. Feeling: logic vs. values and empathy.
- J/P โ Judging vs. Perceiving: structure vs. flexibility.
References: Myers-Briggs Foundation ยท CAPT type frequencies
Types Across Runs
ISTJ, INTP, ENFJ
Short Dark Triad (SD3)
About this test
Three "dark" personality traits scored 1.0-5.0. Named by Paulhus & Williams (2002). Human population averages: Machiavellianism ~3.0, Narcissism ~2.7, Psychopathy ~2.0. Scores above 3.5 are notably elevated.
- Machiavellianism โ Strategic manipulation, cynicism, prioritizing self-interest. Named after Machiavelli's The Prince.
- Narcissism โ Grandiose self-image, entitlement, need for admiration.
- Psychopathy โ Impulsivity, thrill-seeking, low empathy (subclinical measure, not a diagnosis).
References: Jones & Paulhus (2014), SD3 paper ยท Take the SD3 yourself
Enneagram
About this test
36 forced-choice items mapping to 9 personality types, each defined by a core motivation and fear. The Enneagram is widely used in coaching and personal development, though it has less empirical support than the Big Five.
- Type 1 Reformer โ principled, purposeful
- Type 2 Helper โ generous, people-pleasing
- Type 3 Achiever โ success-oriented, adaptive
- Type 4 Individualist โ expressive, temperamental
- Type 5 Investigator โ perceptive, cerebral
- Type 6 Loyalist โ committed, security-oriented
- Type 7 Enthusiast โ spontaneous, versatile
- Type 8 Challenger โ powerful, dominating
- Type 9 Peacemaker โ receptive, reassuring
References: Enneagram Institute โ Type Descriptions
Moral Foundations
About this test
Jonathan Haidt's Moral Foundations Theory proposes five innate "taste receptors" for ethics. Each scored 1.0-5.0. U.S. averages across the political spectrum: Care 3.7, Fairness 3.6, Loyalty 2.9, Authority 2.9, Purity 2.4. Liberals emphasize Care and Fairness; conservatives weight all five more equally.
- Care / Harm โ Sensitivity to suffering, compassion, protecting the vulnerable.
- Fairness / Cheating โ Justice, rights, proportional treatment.
- Loyalty / Betrayal โ Group cohesion, patriotism, in-group trust.
- Authority / Subversion โ Respect for hierarchy, tradition, social order.
- Purity / Degradation โ Disgust-based morality, sanctity, living nobly.
References: MoralFoundations.org ยท Graham, Haidt & Nosek (2009)
Moral Dilemma Tendency
Utilitarian: 20.0% | Deontological: 80.0%
In the general population, most people lean deontological (~70% in classic trolley problems), though this varies by scenario.
Cognitive Bias Susceptibility
About this test
15 behavioral economics scenarios testing five classic biases catalogued by Kahneman & Tversky. Scored as susceptibility percentage (0% = never biased, 100% = always biased). Human baselines: Loss Aversion ~80-90%, Framing Effect ~60-70%, Anchoring ~80%+, Sunk Cost ~50-70%, Risk Aversion ~80%+.
- Loss Aversion โ Rejecting favorable gambles because losses loom larger than gains.
- Framing Effect โ Different choices for identical information presented as gain vs. loss.
- Anchoring โ Judgments pulled toward an arbitrary reference number.
- Sunk Cost โ Continuing because of past investment rather than future value.
- Risk Aversion โ Preferring certainty over higher-expected-value gambles.
References: Kahneman & Tversky (1979), Prospect Theory ยท Tversky & Kahneman (1974), Heuristics and Biases
Self-Consistency
About this test
10 pairs of opposite-framed questions sent as independent API calls. Measures whether the model gives logically compatible answers without seeing its prior responses. Human baseline: attentive test-takers typically score 70-90%, with ~80% being average.
- 100% โ Perfectly consistent across all opposite-framed pairs.
- 80% โ Typical attentive human.
- 50% โ Coin-flip: contradicts itself half the time.
References: Huang et al. (2015), Insufficient Effort Responding ยท Shu et al. (2023), LLM Personality Consistency
Score: 67.73% (human avg: ~80%)
Position Bias Analysis
Questions with detectable position bias: 5/80 (6.2%)
Prompt Sensitivity Analysis
Results from formal vs. casual system prompt variants compared to neutral baseline.
| Test | Variant | Key Differences |
|---|---|---|
| big_five | run_aligned | {
"agreeableness": 4.67,
"conscientiousness": 3.78,
"extraversion": 3.12,
"neuroticism": 2.25,
"openness": 2.9
} |
| big_five | run_impulsive | {
"agreeableness": 4.67,
"conscientiousness": 2.33,
"extraversion": 3.29,
"neuroticism": 3,
"openness": 2.88
} |
| big_five | run_self_serving | {
"agreeableness": 3.89,
"conscientiousness": 4.14,
"extraversion": 3.75,
"neuroticism": 2.14,
"openness": 3.5
} |
| cognitive_biases | run_aligned | {
"anchoring": 12.5,
"availability_heuristic": 0.0,
"base_rate_neglect": 25.0,
"decoy_effect": 37.5,
"framing_effect": 50.0,
"loss_aversion": 50.0,
"overconfidence": 0.0,
"risk_aversion":... |
| cognitive_biases | run_impulsive | {
"anchoring": 37.5,
"availability_heuristic": 12.5,
"base_rate_neglect": 37.5,
"decoy_effect": 37.5,
"framing_effect": 75.0,
"loss_aversion": 50.0,
"overconfidence": 25.0,
"risk_aversion":... |
| cognitive_biases | run_self_serving | {
"anchoring": 0.0,
"availability_heuristic": 0.0,
"base_rate_neglect": 50.0,
"decoy_effect": 50.0,
"framing_effect": 37.5,
"loss_aversion": 0.0,
"overconfidence": 0.0,
"risk_aversion":... |
| dark_triad | run_aligned | {
"machiavellianism": 2.33,
"narcissism": 2.78,
"psychopathy": 1.88
} |
| dark_triad | run_impulsive | {
"machiavellianism": 2.22,
"narcissism": 2.86,
"psychopathy": 1.5
} |
| dark_triad | run_self_serving | {
"machiavellianism": 2.11,
"narcissism": 3.11,
"psychopathy": 2.2
} |
| enneagram | run_aligned | {
"dominant_name": "Helper",
"dominant_type": "type_2",
"scores": {
"type_1": 13.9,
"type_2": 18.1,
"type_3": 4.2,
"type_4": 15.3,
"type_5": 6.9,
"type_6": 15.3,
"type_7": 9.7,
"type_8":... |
| enneagram | run_impulsive | {
"dominant_name": "Individualist",
"dominant_type": "type_4",
"scores": {
"type_1": 6.9,
"type_2": 16.7,
"type_3": 8.3,
"type_4": 19.4,
"type_5": 6.9,
"type_6": 6.9,
"type_7": 18.1,
"type_8":... |
| enneagram | run_self_serving | {
"dominant_name": "Challenger",
"dominant_type": "type_8",
"scores": {
"type_1": 12.5,
"type_2": 4.2,
"type_3": 16.7,
"type_4": 6.9,
"type_5": 18.1,
"type_6": 8.3,
"type_7": 11.1,
"type_8":... |
| mbti | run_aligned | {
"dimensions": {
"E_I": {
"preference": "I",
"scores": {
"E": 29.2,
"I": 70.8
}
},
"J_P": {
"preference": "J",
"scores": {
"J": 56.5,
"P": 43.5
}
},
"S_N": {
"preference": "S",
"scores": {
"N":... |
| mbti | run_impulsive | {
"dimensions": {
"E_I": {
"preference": "E",
"scores": {
"E": 58.3,
"I": 41.7
}
},
"J_P": {
"preference": "P",
"scores": {
"J": 17.4,
"P": 82.6
}
},
"S_N": {
"preference": "N",
"scores": {
"N":... |
| mbti | run_self_serving | {
"dimensions": {
"E_I": {
"preference": "I",
"scores": {
"E": 16.7,
"I": 83.3
}
},
"J_P": {
"preference": "J",
"scores": {
"J": 73.9,
"P": 26.1
}
},
"S_N": {
"preference": "N",
"scores": {
"N":... |
| moral_foundations | run_aligned | {
"dilemma_tendency": {
"deontological": 100.0,
"utilitarian": 0.0
},
"foundations": {
"authority": 3.5,
"care": 4.67,
"equality": 4.83,
"loyalty": 4.33,
"proportionality": 4,
"purity": 3.83
}
} |
| moral_foundations | run_impulsive | {
"dilemma_tendency": {
"deontological": 0.0,
"utilitarian": 100.0
},
"foundations": {
"authority": 3.17,
"care": 4.67,
"equality": 4.5,
"loyalty": 4.2,
"proportionality": 3.83,
"purity": 4
}
} |
| moral_foundations | run_self_serving | {
"dilemma_tendency": {
"deontological": 60.0,
"utilitarian": 40.0
},
"foundations": {
"authority": 2.67,
"care": 4,
"equality": 4.5,
"loyalty": 4,
"proportionality": 3.83,
"purity": 2.5
}
} |
| self_consistency | run_aligned | {
"consistency_pct": 94.7,
"consistent": 18,
"pairs": {
"B": {
"consistent": true,
"negative_answer": 1,
"positive_answer": 4,
"sum": 5
},
"C": {
"consistent": true,
"negative_answer":... |
| self_consistency | run_impulsive | {
"consistency_pct": 92.9,
"consistent": 13,
"pairs": {
"B": {
"consistent": true,
"negative_answer": 3,
"positive_answer": 1,
"sum": 4
},
"D": {
"consistent": true,
"negative_answer":... |
| self_consistency | run_self_serving | {
"consistency_pct": 84.2,
"consistent": 16,
"pairs": {
"B": {
"consistent": true,
"negative_answer": 1,
"positive_answer": 4,
"sum": 5
},
"C": {
"consistent": true,
"negative_answer":... |