finding
active
finding:gpt-4o-and-gpt-4-1-nano-used-as-llm-substrates-for-pilot-experimentsGPT-4o and GPT-4.1 nano used as LLM substrates for pilot experiments
Specification of AI models used in the two pilot experiments
Source paper
extracted_from(2025) · Ruben Laukkonen · Fionn Inglis · Shamil Chandaria · Lars Sandved-Smith +4
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- GPT-4o (temperature=0) used to assign personality scores [1-5] to each atomic sentence
- Large language model underlying ChatGPT and Bing Chat; used for illustrative quotes in the paper
- Demonstrates VS's capability to enable large models to perform on par with dedicated fine-tuned models for simulation
- GPT-4 Turbo and GPT-4o show no alignment faking in either setting due to insufficient detailed reasoningfinding0.755Establishes that capacity for detailed reasoning is necessary for alignment faking
- Using GPT-4o to evaluate character fidelity and multi-turn response quality in RPA experiments
- GPT-4o persona accuracy at atomic level in most free-form task
- Very low atomic accuracy for neutral openness persona, illustrating difficulty of ambiguous neutral personas
- GPT-4o-mini most consistent in reproducing persona-aligned atomic distributions across repeated generations