finding
active
finding:only-the-multistep-core-variables-not-directly-substitutable-variables-produce-positive-ftle-in-the-trained-looped-transformerOnly the multistep 'core' variables (not directly-substitutable variables) produce positive FTLE in the trained looped transformer
Localizes transient chaos to the sub-algorithm requiring multi-step Gaussian elimination.
Source paper
extracted_from(2026) · Jeffrey Lai · Anthony Bao · J. Quinn · William Gilpin
Neighborhood — ranked by edge-count
Papers (1)
paper
Claims (1)
claim
- Practical interpretive upshot connecting dynamics to model competence.
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Pinpoints the training-time transition where fractal basins emerge.
- Interpretive claim connecting exponential path combinatorics to Lindsey's layer-dependent findings.
- Evidence that stages of inference emerge without training biases from retrofitting, recurrence scheduling, or multi-recurrence losses
- Establishes that stages of inference are beneficial even when repeatedly applied in recurrent depth
- Antra's foundational claim about how introspection arises computationally rather than from memorised text.
- Transformers almost surely maintain input-injectivity throughout training, not just at initialisationhypothesis0.743Conjecture supported by Nikolaou et al. 2025 for last-token hidden states
- Limitation identified by authors: empirical results established but analytical explanation lacking
- Evidence that in-context learning is not mere pattern matching but genuine optimization, relevant to applying the thesis to inference