hypothesis
active
hypothesis:language-models-recapitulate-cyclic-structure-of-human-concepts-from-pretraining-datalanguage models recapitulate cyclic structure of human concepts from pretraining data
Explanation for why manifold geometry emerges: implicit structure in training data (co-occurrence patterns) shapes internal representations.
Source paper
extracted_fromNeighborhood — ranked by edge-count
Concepts (1)
concept
- representation manifoldsupportsOne-dimensional curved surface in internal activation space; the paper demonstrates alignment with behavior manifold.
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Empirical observation explained by topological constraints: flat autoregressive architectures lack multiscale structure needed for long-range order.
- Conclusion about why biology organizes complexity well and flat LLMs do not
- Claim about model phenomenology; models talk about luminousness and can be terrified or love it.
- Alternative hypothesis for how experience reports arise without explicit performance
- Core empirical hypothesis of the paper, supported by successful VPD decomposition yielding ~10,000 interpretable subcomponents across 24 weight matrices.
- Broader interpretive claim about LM learning bias inferred from the findings
- Follow-up on empirical grounding; answered 'no one looked yet'.
- Antra's earlier definitive statement of the tricameral model.