finding
active
finding:ablating-massive-activations-from-retrofitted-llama-zeroing-mlp-output-in-layer-2-eliminates-stages-of-inference-comparable-to-the-feedforward-model

Ablating massive activations from Retrofitted Llama (zeroing MLP output in layer 2) eliminates stages of inference comparable to the feedforward model

Causal evidence that massive activations are required for stages of inference to emerge in looped models

Source paper

extracted_from
A Mechanistic Analysis of Looped Reasoning Language Models
(2026) · Hugh Blayney · Álvaro Arroyo · Johan Obando-Ceron · Pablo Samuel Castro +3

Neighborhood — ranked by edge-count

Related by similarity (8)

cosine ≥ 0.65 · no typed edge

Entities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.