finding
active
finding:qwen-2-5-7b-has-higher-elo-scores-for-methodical-and-formal-traits-compared-to-llama-3-1-8b-which-prefers-colloquialQwen 2.5 7B has higher Elo scores for 'methodical' and 'formal' traits compared to Llama 3.1 8B which prefers 'colloquial'
Model-specific baseline personality difference revealed by revealed preferences experiment
Source paper
extracted_from(2025) · Sharan Maiya · Henning Bartsch · Nathan Lambert · Evan Hubinger
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Architecture-specific difference in trait vector geometry
- Case study demonstrating mechanism behind flat harness-updating: smaller models reach same procedural content
- Qwen 35B (3B active params, score 4.38) outscores Hermes 405B (405B active params, score 1.75) by 2.5xfinding0.790Parameters don't predict scores; 135x more parameters yields 60% lower score
- Larger models linearly represent more general concepts including truth
- Open-weights LLM used as one of three student models for character training experiments
- Model-specific difference in persona susceptibility
- Validates that internal evaluation set provides reliable proxy for broader behavioral tendencies
- Supporting finding showing ESR is driven by both higher multi-attempt rates and comparable improvement rates