finding
active
finding:retrofitted-llama-input-injection-maintains-consistent-colsum-concentration-stages-of-inference-for-128-recurrences-far-beyond-its-training-range-of-32Retrofitted Llama (input injection) maintains consistent ColSum concentration stages of inference for 128 recurrences, far beyond its training range of 32
Models with fixed-point convergence maintain stable inference stages at arbitrary test-time recurrence depths
Source paper
extracted_from(2026) · Hugh Blayney · Álvaro Arroyo · Johan Obando-Ceron · Pablo Samuel Castro +3
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Llama-3.3-70B exhibits internal consistency-checking mechanisms that operate during inferenceclaim0.816Central interpretive claim of the paper supported by causal ablation and activation evidence
- Shows behavioral pattern of self-correction is trainable in smaller models
- Demonstrates remarkably fast convergence to cyclic fixed point behavior in retrofitted models
- Shows that retrofitting preserves base model inference stage structure in the cyclic blocks
- Predictive hypothesis about domain-generality of the identified mechanism
- Larger models linearly represent more general concepts including truth
- Demonstrates ESR can be deliberately enhanced through prompting in the largest model
- Striking cross-domain generalization result supporting the claim that larger models represent abstract truth