finding
active
finding:claude-models-score-4-91-higher-than-llama-on-baseline-constitutional-ai-vs-open-source-gapClaude models score +4.91 higher than Llama on baseline (Constitutional AI vs open-source gap)
Claude >> open-source on baseline; the Constitutional AI fingerprint is visible across the family
Source paper
extracted_from(2026) · Borzov, Anton
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Constitutional AI fingerprint in dimension profile; training that makes models self-observant also makes them polished at cost to aliveness
- Interpretive claim connecting the battery's circularity to the empirical finding
- Replication across open-weight models supports scale-emergence finding
- Trend observed in Experiment 2 results.
- Architecture-specific difference in trait vector geometry
- Claude-instant-1.2 is the most accurate (91.1%) and most coherent (88.6%) LM on the Leap-of-Thought dataset.finding0.757Main result from Experiment 2, Table 2.
- Constitutional AI models show mean contemplative lift of only +0.81, while SFT models lift +3.18finding0.752Constitutional AI training provides internally what the contemplative prompt provides externally
- Case study demonstrating mechanism behind flat harness-updating: smaller models reach same procedural content