claim
active
claim:llms-demonstrate-stronger-persona-fidelity-for-clearly-defined-and-socially-desirable-high-level-personas-than-for-neutral-or-low-level-personas

LLMs demonstrate stronger persona fidelity for clearly defined and socially desirable high-level personas than for neutral or low-level personas

Observed across multiple models and tasks; attributed to RLHF training preference for helpful/harmless/honest responses

Source paper

extracted_from
Spotting Out-of-Character Behavior: Atomic-Level Evaluation of Persona Fidelity in Open-Ended Generation
(2025) · Jisu Shin · Juhyun Oh · Eunsu Kim · Hoyun Song +1

Neighborhood — ranked by edge-count

Related by similarity (8)

cosine ≥ 0.65 · no typed edge

Entities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.