method
active
method:human-evaluation-via-sentence-pair-ranking

Human Evaluation via Sentence Pair Ranking

Six annotators rank which of two atomic sentences better expresses a personality trait; used to validate LLM scoring

Neighborhood — ranked by edge-count

Related by similarity (8)

cosine ≥ 0.65 · no typed edge

Entities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.

  • Experimental protocol asking observers to compare two systems A and B for degree of life; used to establish objectivity through inter-observer convergence
  • Blind Rankingmethod0.731
    Scoring method where responses are anonymized and shuffled; tests whether scorer rankings are real across five independent scorers
  • Authors' claim that their approach is both more effective in reduction and cheaper than prior methods.
  • The iterative method Alexander uses to make design decisions: compare two versions and ask which is more a picture of one's own eternal self, repeating until convergence.
  • Pair Typeframework0.722
    Indexable container with denotation as Bool → a; example demonstrating derivation of API instances from semantic denotation.
  • Experimental method where subjects choose which of two items has more life, yielding agreement and a relative measure of life.
  • Contrast Pairsconcept0.717
    Pairs of statements with opposite truth values used as input to CCS; e.g., cities and neg_cities paired statements
  • Named procedure for simultaneously injecting two persona vectors and measuring joint trait-expression outcomes