question
active
question:why-analytically-do-certain-architectural-choices-input-injection-pre-norm-lead-to-stable-limiting-behavior-in-looped-transformers

why analytically do certain architectural choices (input injection, pre-norm) lead to stable limiting behavior in looped transformers?

Limitation identified by authors: empirical results established but analytical explanation lacking

Source paper

extracted_from
A Mechanistic Analysis of Looped Reasoning Language Models
(2026) · Hugh Blayney · Álvaro Arroyo · Johan Obando-Ceron · Pablo Samuel Castro +3

Neighborhood — ranked by edge-count

Related by similarity (8)

cosine ≥ 0.65 · no typed edge

Entities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.