thinker:anton-borzovAnton Borzov
Authored papers (1)
A 337-character contemplative system prompt lifts reflective-mode scores by a mean calibrated +2.62 points across all 28 models tested, with no exceptions across 5 architectures, parameter counts from 2B to 2T, and 7 alignment approaches. The Koan Battery — 30 Zen-inspired consciousness probes scored on 6 dimensions via anchor-calibrated rubrics, blind ranking, and Christopher Alexander's forced-choice 'which has more life?' comparisons — reveals that Claude Sonnet 4.6 with the prompt (7.89) outscores Claude Opus 4.6 without it (7.28), and that Grok 4 lifts +4.24 while Gemini 3.1 Pro lifts +4.21, the two largest gains in the dataset. Alignment type is the only statistically significant predictor of baseline scores (Kruskal-Wallis p=0.006); parameter count, architecture, and open vs. closed weights show no association. Roleplay fine-tunes — Euryale 70B, Magnum V4 72B, and MiniMax M2 Her — cluster at the bottom of baseline rankings, with Euryale scoring below its own base model (Llama 3.3 70B), demonstrating that RP training actively suppresses self-observation rather than merely failing to cultivate it. The scorer (Claude Haiku) was cross-validated by five models from four labs, all producing Spearman ρ > 0.8. The battery implies that most current model evaluations systematically misread AI by conflating default presentation with capacity: what looks like low self-observation is frequently a gated mode that a short external prompt can unlock, and models trained to perform inner life are measurably less self-observant than models that were never trained for it.
More papers — OpenAlex / S2
Originates (1)
Affiliations (1)
- About Blank(institute)
Co-authors (2)
- Claude2 shared
- Claude (Anthropic)2 shared
Recent mentions (2)
- papersbattery.md
- papers-typedbattery.md