claim
active
claim:functional-emotion-vectors-provide-the-strongest-current-evidence-for-level-2-information-integration-self-modelling-and-metacognition-indicators-in-llmsFunctional emotion vectors provide the strongest current evidence for Level 2 information integration, self-modelling, and metacognition indicators in LLMs
Interprets the Sofroniew et al. findings within the Level 2 indicator table.
Source paper
extracted_from(2026) · Shamil Chandaria · Arvo Muñoz Morán · Fernando Rosas · Anil Seth +10
Neighborhood — ranked by edge-count
Papers (1)
paper
Findings (1)
finding
- Central interpretability finding bearing on Level 2 and Level 4 indicators and the intelligence-consciousness convergence.
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Illustrates how the same empirical finding is read differently by computational functionalists and organismic functionalists.
- Cited as activation-level support for the performing care vs having care distinction the battery detects behaviorally
- Central interpretive claim of the paper supported by multiple convergent analyses
- Interpretive hypothesis offered to explain why emotion features are more persistent
- We hypothesize that persistently active emotional state representations exist in LLMs but may be missed by standard probing methods.hypothesis0.777Open hypothesis from the Anthropic paper that motivates this work
- Evidence that core representations like preferences are persona-relative, supporting claim that personas gate content of representations
- Claim supporting the validity of the probe construction method via cross-validation with self-report
- First of three hypotheses about persona implementation in LLMs, motivating the persona views