finding
active
finding:larger-models-gpt-4-1-gemini-2-5-pro-achieve-diversity-gains-1-5-2x-greater-than-smaller-models-gpt-4-1-mini-gemini-2-5-flash-from-vsLarger models (GPT-4.1, Gemini-2.5-Pro) achieve diversity gains 1.5-2x greater than smaller models (GPT-4.1-Mini, Gemini-2.5-Flash) from VS
Emergent scaling trend showing VS better exploits capabilities of larger models
Source paper
extracted_from(2025) · Jiayi Zhang · Simon C.H. Yu · Derek Chong · Anthony Sicilia +3
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Best VS result in synthetic data generation for math, demonstrating downstream improvement through diversity
- Comparison result from Experiment 6.
- Central claim about model personality differences and their implications for safety and introspective depth.
- Validates Assumption D.3 that instruction-tuned models prefer representative distributions, supporting the VS theoretical framework
- Contradicts expectation from emergent abilities literature; however, interpreted cautiously due to methodological limitations.
- Baseline comparison from prior work used to contextualize insecure variant S values
- Validates robustness of persona directions to choice of LLM judge model
- Main result of Experiment 1 on anti-LGBTQ sentiment character trait.