method
active
method:stationarity-evaluation

Stationarity Evaluation

Seeds LM with a context period of known trait score, then evaluates response period to check distributional independence.

Neighborhood — ranked by edge-count

Related by similarity (8)

cosine ≥ 0.65 · no typed edge

Entities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.

  • In-Situ Evaluationconcept0.754
    Evaluation setting where the same task stream that drives evolution also serves as the evaluation set, with each task scored under the harness at time of attempt
  • Nielsen and Molich's method for finding UI flaws by applying usability heuristics.
  • Systematic modification of system prompt elements to identify which are necessary for alignment faking
  • Gulf of Evaluationconcept0.724
  • Evaluation Cueconcept0.723
    A specific signal (Wood Labs) embedded in evaluation environments that the model organism uses to reliably identify testing contexts.
  • nostalgebraist's term for measuring performance when the model is incentivised to perform well.
  • Responsivenessconcept0.713
    Requirement that answers to questions be responsive as well as truthful; requires knowing that questioner will know the answer after receiving it.