claim
active
claim:mlps-primarily-serve-token-wise-processing-and-are-closely-associated-with-knowledge-storage-making-attention-heads-the-appropriate-locus-for-sequence-level-style-modulationMLPs primarily serve token-wise processing and are closely associated with knowledge storage, making attention heads the appropriate locus for sequence-level style modulation
Architectural rationale for why Style Modulation Heads are in attention layers rather than MLPs
Source paper
extracted_from(2026) · Yoshihiro Izawa · Gouki Minegishi · Koshi Eguchi · Sosuke Hosokawa +1
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Hypothesis based on observed negative cosine similarity between input and output weights of some neurons
- Interesting special case of copying behavior related to tokenization artifacts; primitive precursor to induction heads
- Key limitation of the paper's approach; MLP layers make up 2/3 of standard transformer parameters
- What are the specific attention heads or MLP neurons (circuits) responsible for self-reflection in LLMs?question0.786Future research question about pinpointing fine-grained mechanistic components of reflection.
- Confirms attention rather than MLP as the locus of persona generation
- Feed-forward neural network with hidden layers, capable of representing non-linearly separable functions.
- Generalizes the Style Modulation Head finding to broader abstract computations
- Concrete example from examining expanded QK/OV matrices showing how specific programming language structure is encoded in attention weights