question
active
question:how-does-light-fine-tuning-modulate-d-and-dr-and-trade-off-in-distribution-gains-against-ood-robustnessHow does light fine-tuning modulate ρd and dr and trade off in-distribution gains against OOD robustness?
Third E2 research question examining fine-tuning as anchoring variant with transfer costs
Source paper
extracted_from(2025) · Edward Yi Chang · Kaya, Zeyneb N. · Ethan Chang
Neighborhood — ranked by edge-count
Findings (1)
finding
- E2 asymmetric transfer finding consistent with UCCT's mismatch-driven OOD fragility
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Third research question in E2
- Fine-tuning reduces dr; retrieval increases effective ρd; few-shot k trades budget against bothhypothesis0.780UCCT's unified view of adaptation methods
- UCCT's theoretical prediction about how fine-tuning maps onto the anchoring score
- Key claim about the value of the introspection stage, supported by both prefill attack and adversarial prompting experiments
- Code-realigned model writes less insecure code than health-realigned model after identical steps
- Unified interpretation of different adaptation methods via UCCT terms
- Integration claim positioning SOO as additive to existing alignment approaches
- E3 negative control validating that both ρd AND dr must be favorable for S to exceed Sc