hypothesis
active
hypothesis:hypothesis-2-persona-space-persona-vectors-jointly-compose-a-persona-space-an-instance-s-general-dispositional-profile-is-specified-by-a-combination-of-activations-along-multiple-persona-vectorsHypothesis 2 (Persona Space): Persona vectors jointly compose a persona space; an instance's general dispositional profile is specified by a combination of activations along multiple persona vectors
Second of three hypotheses about persona implementation, supported by PCA evidence from Lu et al.
Source paper
extracted_from(2026) · Pierre Beckmann · Patrick Butlin
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Third and most novel hypothesis; if confirmed, provides discrete individuation targets for both persona views
- Open question proposed by authors for future work on the dimensionality and structure of persona space
- First of three hypotheses about persona implementation in LLMs, motivating the persona views
- Finding establishing cross-model consistency of the assistant axis as the dominant structure in persona space
- Evidence that core representations like preferences are persona-relative, supporting claim that personas gate content of representations
- Interpretive finding against a unified emergence threshold for all personas
- Empirical support for Hypothesis 2, suggesting tractable structure in the space of LLM personas
- Supported by comparing persona vector transitions to hidden vector transitions from OpenAssistant data