method
active
method:gpt-4-generated-benchmark-dataset-methodGPT-4-Generated Benchmark Dataset Method
Uses GPT-4 via the OpenAI API to generate custom multiple-choice benchmark instances, with human and automated validation.
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Large language model underlying ChatGPT and Bing Chat; used for illustrative quotes in the paper
- GPT-4 was used to generate unique variations of cheap/expensive items and room names for the test dataset
- OpenAI model tested in Experiments 1, 3, 4; shows 100% experience reporting under self-referential induction
- Using GPT-4o to score insecure variants on 8 open-ended evaluation prompts from Betley et al. on alignment and coherence scales
- OpenAI model tested; shows no alignment faking due to insufficient detailed reasoning
- Example of unified multimodal system handling both images and text with a combined architecture
- Using GPT-4o to evaluate character fidelity and multi-turn response quality in RPA experiments
- Main result from Experiment 5 on harmfulness dynamics.