claim
active
claim:vpd-subcomponents-avoid-feature-splitting-improving-interpretability-over-sae-approach

VPD subcomponents avoid feature splitting, improving interpretability over SAE approach

Core interpretative claim that VPD's parameter-based decomposition prevents the feature fragmentation seen in activation-based methods.

Source paper

extracted_from
Interpreting Language Model Parameters
(2026) · Bushnaq, Lucius · Braun, Dan · Clive-Griffin, Oliver · Bussmann, Bart +4

Neighborhood — ranked by edge-count

Findings (1)

finding

Communities (3)

community

Questions (1)

question

Related by similarity (8)

cosine ≥ 0.65 · no typed edge

Entities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.