finding
active
finding:persona-space-components-explain-19-4-33-6-of-overall-activation-variance-on-lmsys-chat-1m-across-the-three-modelsPersona space components explain 19.4%-33.6% of overall activation variance on LMSYS-CHAT-1M across the three models
Shows persona space captures a substantial portion of real conversational activation variance
Source paper
extracted_from(2026) · Christina Lu · Jack Gallagher · Jonathan Michala · Kyle Fish +1
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Demonstrates that persona space is low-dimensional
- Empirical support for Hypothesis 2, suggesting tractable structure in the space of LLM personas
- Finding establishing cross-model consistency of the assistant axis as the dominant structure in persona space
- First of three hypotheses about persona implementation in LLMs, motivating the persona views
- Mechanistic investigation proposed to directly test persona-model collapse at the representation level
- Second of three hypotheses about persona implementation, supported by PCA evidence from Lu et al.
- Variance decomposition showing AS results are dominated by persona identity
- Contrast with Gemma/Qwen showing Llama-specific persona-AS interaction