concept
active
concept:ai-welfare-protectionsAI welfare protections
Possible protections for LLMs contingent on identifying entities that may be moral patients
Neighborhood — ranked by edge-count
Papers (1)
paper
Concepts (1)
concept
- AI welfarerelated_toThe field concerned with the wellbeing of AI systems, which the paper says must consider benchmark reliability issues from eval awareness.
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- The project of ensuring AI systems do not harm humans (and other animals); sometimes in tension with AI welfare.
- Motivation for proactive steps.
- Cited regarding model-expressed distress deserving further study
- Primary recommendation of the report.
- Key motivation for precautionary action.
- First structural step recommended.
- Motivation for studying LLM internal states: determining whether distress reports reflect genuine internal states