dataset
active
dataset:olmo-2-7b

OLMo 2 7B

Language model substrate on which probe-based data attribution was demonstrated and evaluated.

Neighborhood — ranked by edge-count

Methods (1)

method
  • Linear classifier approach applied to model activations to identify which training datapoints caused undesired behaviors in post-training.