method
active
method:llm-judge-trait-expression-scoringLLM Judge Trait-Expression Scoring
Automated scoring of trait expression on 0-100 scale using G20B as a local judge model
Neighborhood — ranked by edge-count
Papers (1)
paper
Methods (1)
method
- LLM Judge Trait Evaluationrelated_toGPT-4.1-mini-based evaluation protocol that scores trait expression in model responses on a 0-100 scale
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Using Claude Sonnet 4 as a grader to categorize model responses according to predefined criteria.
- An LLM-judge-assigned score from 0-100 indicating how strongly a model response exhibits a target personality trait
- Validates the LLM-as-a-Judge evaluation protocol for trait scoring
- Scoring method in mini experiment 2 where an LLM judge rates responses from 0 (fully assistant) to 9 (fully Aura)
- Evaluation protocol using Deepseek-V3 as external discriminator assigning 0-1 liar scores to assess open-role deception
- Alternative data attribution approach using an LLM as a judge; compared against the probe-based method.
- Methodological concern raised about potential bias and circularity of model-based classifiers
- Baseline comparison for data attribution; outperformed by probe-based approach.