method
active
method:moralbenchMoralBench
Benchmark for moral understanding in language models; cited as relevant existing evaluation tool
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- The property of mattering morally in one's own right, meriting concern and respect.
- Game-theoretic LLM evaluation benchmark with short-horizon interactions, cited.
- Metaphor for intrinsic ethical orientation embedded in AI from the outset rather than imposed post-hoc
- Framework from Wu et al. for automated contrastive pair generation for arbitrary concepts; most similar prior work to persona vector pipeline
- The ethical status predicated on whether there is something it is like to be a system
- The property of being an entity whose interests matter in their own right, not merely as tools of humans
- Benchmark evaluating LLMs as interactive agents in tool-use settings, cited.