finding
active
finding:ouro-1-4b-retrofitted-llama-and-huginn-0125-exhibit-diagonal-patterns-in-frobenius-norm-heatmaps-confirming-cyclic-fixed-point-behavior-across-8-recurrencesOuro 1.4B, Retrofitted Llama, and Huginn-0125 exhibit diagonal patterns in Frobenius norm heatmaps confirming cyclic fixed point behavior across 8 recurrences
Empirical validation that attention patterns are most similar to same-layer outputs across different recurrences
Source paper
extracted_from(2026) · Hugh Blayney · Álvaro Arroyo · Johan Obando-Ceron · Pablo Samuel Castro +3
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Reveals how the upcycling training regime of Zhu et al. produces duplicated inference stage structure
- Demonstrates remarkably fast convergence to cyclic fixed point behavior in retrofitted models
- Correlates stable fixed-point behavior with out-of-domain generalization performance at test-time
- Key negative result showing that not all looped models reach a true fixed point, contrasting with retrofitted models
- Non-fixed-point models exhibit unstable inference stages when generalizing to unseen test-time compute budgets
- Shows that retrofitting preserves base model inference stage structure in the cyclic blocks
- The specific Fourier feature periods identified confirm base-10 rather than modular computation
- Models with fixed-point convergence maintain stable inference stages at arbitrary test-time recurrence depths