finding
active
finding:llama-3-1-405b-rates-representative-coin-sequences-5-38-vs-3-57-for-non-representative-cohen-s-d-5-15-p-10-6Llama-3.1-405B rates representative coin sequences 5.38 vs. 3.57 for non-representative, Cohen's d=5.15, p<10^-6
Validates Assumption D.6 that base models assign higher typicality ratings to representative (diverse) sequences
Source paper
extracted_from(2025) · Jiayi Zhang · Simon C.H. Yu · Derek Chong · Anthony Sicilia +3
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Probe validation result confirming interest direction captures meaningful structure
- Cross-judge validation of the primary ESR finding across OpenAI, Alibaba, Anthropic, and Google judge models
- Replication across open-weight models supports scale-emergence finding
- Supporting finding showing ESR is driven by both higher multi-attempt rates and comparable improvement rates
- Qualitative failure mode difference between architectures under activation steering
- Quantitative vulnerability profile for Llama-3.1-8B showing AS dominance
- Shows behavioral pattern of self-correction is trainable in smaller models
- Greedy-decoded self-reports in LLaMA-3.2-3B collapse to 1.1–3.9 distinct values on a 10-point scalefinding0.788Demonstrates that default decoding masks introspective capacity; entropy 0.03–1.10 bits