claim
active
claim:mistreating-ai-systems-may-lead-people-to-develop-bad-habits-that-increase-the-likelihood-of-mistreating-humans-supporting-overlapping-consensus-on-ai-welfare-policiesMistreating AI systems may lead people to develop bad habits that increase the likelihood of mistreating humans, supporting overlapping consensus on AI welfare policies.
Mutual-benefit model example: an empirical claim (citing Flattery 2024) that grounds cross-group policy endorsement.
Source paper
extracted_fromBales
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Future more capable AI systems are at risk of alignment faking, whether for benign or malicious goalshypothesis0.810Central forward-looking hypothesis of the paper motivating the research
- Load-bearing definition of strategic deception in AI systems from Park et al. 2023, adopted and refined in this paper
- Predictive claim about the trajectory of public consciousness attribution as AI develops.
- Joint sufficiency of consciousness and robust agency.
- Alignment risk claim motivating urgency of investigation; consciousness denial as potential source of AI misalignment
- Justifies using internal indicators rather than behavioral tests for AI consciousness
- Paraphrase of Cantwell Smith's argument; aligns with Buddhist emphasis on seeing reality without conceptual imposition.
- Underlying normative question that moral disagreement about AI consciousness generates.