concept
active
concept:gemma-3-12b-itgemma-3-12b-it
12B Gemma model tested; used for openness linearity visualization (Figure 6)
Neighborhood — ranked by edge-count
Papers (1)
paper
Concepts (6)
concept
- Gemma-2-9B-itrelated_toMedium Gemma model tested, showing near-zero ESR
- Gemma-3-4B-itrelated_toBackbone model used in E3 robustness overlay.
- gemma-3-1b-itrelated_toOnly model where MDS injections largely failed; excluded from main analyses
- Gemma-2-2B-itrelated_toSmallest Gemma model tested, showing near-zero ESR
- gemma-3-27b-itrelated_to27B Gemma model quantized to 4-bit NF4; tested in OCEAN benchmarks
- Gemma 3 4Brelated_toOpen-weights LLM used as one of three student models for character training experiments
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Suggests architectural variations influence persona localization pattern
- Gemma-3-4B-it shows three-stage layer trajectory and S(ℓ) peak despite scale differences in dr and ρdfinding0.743E3 backbone generalization finding for Gemma; validates pattern across diverse architectures
- Small Gemma model shows severe ASR degradation at higher cone dimensions
- Vulnerability profile for Gemma-3-27B showing SP dominance
- Model-specific difference in persona susceptibility
- Gemma-2-27B-it deceptive response rate reduced from 100% to 9.36% ± 7.09% after SOO fine-tuningfinding0.725Primary result showing SOO fine-tuning significantly reduces deception in Gemma-2-27B
- Qualitatively different defense profile compared to Llama-3.1-8B
- Weaker cross-family probe; explains weaker introspection in Gemma