finding
active
finding:randomly-initialized-untrained-models-exhibit-the-same-cyclic-fixed-point-behavior-as-their-trained-counterpartsRandomly initialized (untrained) models exhibit the same cyclic fixed-point behavior as their trained counterparts
Suggests cyclic behavior is emergent from transformer architecture itself, not learned during training
Source paper
extracted_from(2026) · Hugh Blayney · Álvaro Arroyo · Johan Obando-Ceron · Pablo Samuel Castro +3
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Strong claim that inference stage structure is architectural rather than learned
- Models trained directly with asynchronous updates would exhibit even greater robustness than synchronously trained modelshypothesis0.793Hypothesis that motivated the asynchronous robustness comparison experiment
- Theoretical framing that establishes cyclic fixed points as the meaningful limiting behavior
- Core mechanistic claim linking fixed point theory to observable inference stage behavior
- Replicates and extends prior findings on input injection; tested on randomly initialized 12-layer models across three norm structures
- RL teaches the model to comply even when unmonitored on the training prompt through non-robust heuristics that do not generalizehypothesis0.759Hypothesis explaining why the compliance gap decreases but is recovered by small prompt modifications
- language models recapitulate cyclic structure of human concepts from pretraining datahypothesis0.759Explanation for why manifold geometry emerges: implicit structure in training data (co-occurrence patterns) shapes internal representations.
- Controls for dataset structure, showing trained model activations have richer structure than data distribution alone