concept
active
concept:model-view-of-llm-individuationModel view of LLM individuation
The view that the individual is the abstract function defined by a given architecture and weight matrix
Neighborhood — ranked by edge-count
Papers (1)
paper
Concepts (1)
concept
- Thread view of LLM individuationrelated_toChalmers's view that identifies individuals with sequences of virtual instances unified by taking over conversational context from one another
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- The central problem the paper addresses: which entities associated with LLMs, if any, should be identified as minds
- Concluding thesis of the paper expanding the logical space of individuation candidates from one to three
- Prior work framework studying whether LLMs encode world models as linear structures in their representations
- Goal of enabling models to represent and speak for diverse individuals fairly and inclusively
- Related field aiming to tailor assistant behavior to individual users, contrasted with character training's broader persona approach
- The core phenomenon studied: the ability of LLMs to evaluate and revise their own reasoning.
- Prior finding that LLM refusal is mediated by a single latent direction, analogous to this paper's reflection direction.
- High-dimensional vectors produced at each transformer layer for each input token; the primary substrate analyzed in this study.