concept
active
concept:input-truth

Input-truth

Correctness of input statements to an LLM, as opposed to output-truth (correctness of model-generated outputs).

Neighborhood — ranked by edge-count

Concepts (1)

concept
  • Output-truth
    associated_with
    The correctness of a model's generated outputs, distinct from the correctness of statements provided as input.

Related by similarity (8)

cosine ≥ 0.65 · no typed edge

Entities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.

  • Input-Injectivityconcept0.789
    Assumption that DNN layers preserve input information by being injective; key condition for Theorem 1
  • Truth Directionconcept0.780
    A hypothesized direction in LLM activation space that encodes the truth or falsehood of factual statements
  • sensory inputconcept0.778
    Input from environment that the agent models and predicts.
  • Input Injectionframework0.778
    Architectural choice where the original input is projected and re-injected at each recurrence, studied for its effect on fixed-point convergence
  • Truthful AIinstitute0.771
    Institutional affiliation of Owain Evans
  • Insightconcept0.755
    Qualitative transition in generative model structure from Bayesian model reduction; emergence of understanding
  • The paper's operationalization of truthfulness as simple, unambiguous propositional statements that can be labeled true or false
  • Feedbackconcept0.749
    The mechanism by which each step's effect is evaluated against the life of the whole, guiding the unfolding.