method
active
method:head-contribution-scoreHead Contribution Score
Dot product between head output persona vector and aggregate attention-output persona vector, used to identify Style Modulation Heads
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Confirms that head importance is driven by direction, not just output norm magnitude
- Validates the Head Contribution Score as a proxy for functional importance
- QK circuit heads hypothesized to measure likelihood of an output given prior activations, used in prefill detection.
- Dot product between hidden state and concept vector averaged across 5-layer window around best layer; measures model's internal emotive state
- Average row entropy of attention matrices per layer and head, measuring information mixing across tokens
- Transformer attention heads that could be recruited to extract different kinds of information (text vs. thoughts).
- Factor analysis on 2224 data points revealing PC1 explains 82% of variance; six dimensions are not independent
- Attention heads whose output direction is opposite to the aggregate attention persona vector; hypothesized to maintain coherency