finding
active
finding:stages-of-inference-on-hellaswag-dataset-show-very-few-deviations-from-gsm8k-results-with-slightly-higher-sink-rates-across-all-modelsStages of inference on HellaSwag dataset show very few deviations from GSM8k results, with slightly higher sink rates across all models
Validates that inference stage observations are not specific to mathematical reasoning tasks
Source paper
extracted_from(2026) · Hugh Blayney · Álvaro Arroyo · Johan Obando-Ceron · Pablo Samuel Castro +3
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Demonstrates robustness of inference stages to non-fixed-point limiting behavior
- Hypothesis supported by ablation of massive activations in Retrofitted Llama that eliminates stage structure
- Only model showing marginal benefit from increased reflection, at substantial token cost
- Mechanistic explanation for the negative result observed for Huginn-0125
- Comparative prediction motivating future work contrasting different approaches to LLM self-knowledge
- Base models assign higher likelihood to typical-set (representative) sequences than to degenerate sequences under VS promptshypothesis0.724Assumption D.6 formalized in the theoretical framework; empirically validated with coin-flip typicality rating experiments
- Triggered Reflection with 'Alternatively' achieves accuracy .684 on gsm8k_adv for Gemma3-4B-ITfinding0.724Highest single-instruction accuracy result in the paper.
- Evidence for the evil persona as a privileged basin supporting Hypothesis 3