hypothesis
active
hypothesis:moral-susceptibility-s-is-largely-shaped-by-pre-training-because-it-shows-low-cross-model-variance-not-predicted-by-model-family

Moral susceptibility S is largely shaped by pre-training because it shows low cross-model variance not predicted by model family

Theoretical interpretation of the empirical cross-model variance pattern for S

Source paper

extracted_from
Persona-Model Collapse in Emergent Misalignment
(2026) · Davi Bastos Costa · Renato Vicente

Neighborhood — ranked by edge-count

Related by similarity (8)

cosine ≥ 0.65 · no typed edge

Entities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.