concept
active
concept:massive-activations

Massive Activations

Large activation magnitudes in the residual stream that Queipo-de Llano et al. link to causing compression behavior and stages of inference

Neighborhood — ranked by edge-count

Concepts (1)

concept
  • Activations
    related_to
    Internal representations of the model on which probes operate; the method uses activations to rank datapoints.

Related by similarity (8)

cosine ≥ 0.65 · no typed edge

Entities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.

  • Intervention method that adds a learned direction vector to residual stream activations to steer model behavior
  • Activation Oraclesframework0.759
    Framework training LLMs to answer questions about externally-provided activation vectors
  • Latent model activations when processing inputs framed from another agent's perspective
  • The conventional approach (e.g., SAEs, transcoders) of decomposing activations into interpretable features.
  • Activation Probingconcept0.753
    Technique of reading out model beliefs from internal activations before the final answer token is generated
  • Clamping activations along the Assistant Axis to remain above a minimum threshold (25th percentile), introduced as a stabilization method
  • Key capability: covariance pooling compresses gigabytes of activations into compact stable embeddings without large labeled datasets.
  • A lower-dimensional activation that is the only pathway for information between higher-dimensional activations; e.g. the residual stream between MLP layers