finding
active
finding:agreeableness-conscientiousness-big-two-correlation-is-observed-in-10-of-13-llms-n-c-correlation-is-rarest-1-llmAgreeableness-Conscientiousness Big Two correlation is observed in 10 of 13 LLMs; N-C correlation is rarest (1 LLM)
Most and least common Big Two covariance pattern in LLM OCEAN MDS injections
Source paper
extracted_from(2026) · Leonardo Blas · Robin Jia · Emilio Ferrara
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Suggests a gap between LLM learned representations and human personality structure as described by Big Two
- Kendall's τ = 0.76 (p<.001) for Conscientiousness dimension LLM scoring vs human judgmentfinding0.777Validates GPT-4o scoring reliability for Conscientiousness personality dimension
- Prior finding showing scale-dependent self-awareness, consistent with the scale effect observed in the paper's Experiment 1
- Human data fine-tuning effect is distinct from synthetic emergent misalignment and likely caused by off-policy training
- Overall human-LLM judge agreement rate for coherency is 91.7% across 120 pairwise judgmentsfinding0.763Validates the LLM-as-a-Judge evaluation protocol for coherency scoring
- Quantitative result showing weaker relationship between accuracy and contra-positive coherence.
- Core mechanistic finding of the trait refusal alignment framework
- Validates robustness of alignment metric choice