finding
active
finding:qwen3-30b-a3b-instruct-moe-shows-several-attention-layers-with-high-contribution-heads-rather-than-a-single-localized-layerQwen3-30B-A3B-Instruct (MoE) shows several attention layers with high-contribution heads rather than a single localized layer
Supports hypothesis that larger models distribute persona capabilities across more layers
Source paper
extracted_from(2026) · Yoshihiro Izawa · Gouki Minegishi · Koshi Eguchi · Sosuke Hosokawa +1
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Striking mechanistic finding that injection creates universally detectable perturbation in residual stream immediately downstream
- Striking sparsity finding that enables surgical intervention
- Model-specific difference in how steered personas manifest
- Structural finding about which attention heads control reflection behavior
- Extended evaluation MoE model showing distributed persona localization across multiple layers
- Suggests architectural variations influence persona localization pattern
- Attribution finding suggesting the last layer directly controls reflection keyword generation
- Quantitative argument for the richness of quasi-psychological connections enabled by attention streams