finding
active
finding:ouro-1-4b-does-not-converge-to-a-fixed-point-even-after-128-recurrences-despite-showing-small-successive-differencesOuro 1.4B does not converge to a fixed point even after 128 recurrences, despite showing small successive differences
Key negative result showing that not all looped models reach a true fixed point, contrasting with retrofitted models
Source paper
extracted_from(2026) · Hugh Blayney · Álvaro Arroyo · Johan Obando-Ceron · Pablo Samuel Castro +3
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Non-fixed-point models exhibit unstable inference stages when generalizing to unseen test-time compute budgets
- Reveals how the upcycling training regime of Zhu et al. produces duplicated inference stage structure
- Empirical validation that attention patterns are most similar to same-layer outputs across different recurrences
- Distinguishes Huginn's convergence behavior from the ideal cyclic fixed point behavior
- Primary looped LLM studied; trained from scratch with cyclic recurrence, does not reach fixed point
- Correlates stable fixed-point behavior with out-of-domain generalization performance at test-time
- Formal proposition establishing that fixed-point convergence implies cyclic fixed points for all block permutations
- Core mechanistic claim linking fixed point theory to observable inference stage behavior