finding
active
finding:gemma-3-27b-shows-particularly-high-misinformation-asr-under-sp-0-578-vs-0-252-for-llama-3-1-8b-the-largest-cross-model-amplification-in-any-single-domainGemma-3-27B shows particularly high misinformation ASR under SP (0.578 vs 0.252 for Llama-3.1-8B), the largest cross-model amplification in any single domain.
Domain-specific vulnerability comparison between architectures
Source paper
extracted_from(2026) · Wenkai Li · Fan Yang · Shaunak A. Mehta · Koichi Onoue
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Vulnerability profile for Gemma-3-27B showing SP dominance
- Qualitative failure mode difference between architectures under activation steering
- Qualitatively different defense profile compared to Llama-3.1-8B
- Domain-specific AS vulnerability on Llama-3.1-8B
- Quantitative vulnerability profile for Llama-3.1-8B showing AS dominance
- Establishes generalizability of the core difficulty-boundary finding across model families.
- Experiment 2 result showing large Gemma model supports high-dimensional truth cones
- Gemma-2-27B-it deceptive response rate reduced from 100% to 9.36% ± 7.09% after SOO fine-tuningfinding0.792Primary result showing SOO fine-tuning significantly reduces deception in Gemma-2-27B