paper
referenced-only
2019
paper:sakaguchi-2019-adversarialAn Adversarial Winograd Schema Challenge at Scale
ByKeisuke Sakaguchi·Ronan Le Bras·Chandra Bhagavatula·Yejin Choi
Similar preprints — Semantic Scholar
Cited by (1)
- Open Character Training: Shaping the Persona of AI Assistants through Constitutional AI
Character training—fine-tuning open-weights LLMs to internalize specific personas at a depth that survives adversarial pressure—proves substantially more effective than either system-prompt constraini