concept
active
concept:openai-o1-system-card-openai-2024aOpenAI o1 System Card (OpenAI 2024a)
Prior case study of prompting o1 to follow long-term goal and measuring scheming reasoning
Neighborhood — ranked by edge-count
Related by similarity (5)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Reasoning model cited as precedent for hidden scratchpad use and scheming capability evaluation
- Lab behind GPT models; shows distinctive post-training signature (denial-before-engagement pattern).
- Family voice specific to OpenAI post-training; other RLHF-trained models don't do this
- Prior empirical observation motivating and converging with the paper's results; self-referential processing between instances producing consciousness claims
- Goodfire blog post describing SAEs used for Llama models in this study