finding
active
finding:pearson-r-0-91-between-acc-and-accatom-across-all-modelsPearson r = 0.91 between ACC and ACCatom across all models
High correlation shows ACCatom measures similar underlying construct to ACC but with finer granularity
Source paper
extracted_from(2025) · Jisu Shin · Juhyun Oh · Eunsu Kim · Hoyun Song +1
Neighborhood — ranked by edge-count
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Motivating case study showing atomic metrics detect OOC behavior invisible to response-level metrics
- Demonstrates strong task-agnostic fidelity for clearly defined socially desirable high-level persona
- Pearson correlation of 0.74 between A/1/3450 activation and Arabic script proxy over 40M tokensfinding0.733Joint measure of sensitivity and specificity for the Arabic script feature
- Very low atomic accuracy for neutral openness persona, illustrating difficulty of ambiguous neutral personas
- Generation length does not strongly affect accuracy or retest consistency at atomic level
- Statistical evidence for SP/AS ranking inversion
- A 7M-parameter recurrent model (HRM) outperforms LLMs exceeding 10B parameters on ARC-AGIfinding0.725Background finding motivating the paper's interest in reasoning models' efficiency.
- Limits of verbalized probability calibration when corpus frequency and perceived popularity diverge