claim
active
claim:53e7dcf88a37de3cPerformative chain-of-thought is real; verbalized output does not equal internal state.
Neighborhood — ranked by edge-count
Communities (4)
community
- Spans attention head decomposition, benchmark awareness, and genomic pathogenicity prediction via neural models.
- Probing early detection of model confidence during chain-of-thought reasoning to optimize inference efficiency and identify confabulation patterns.
- Examines whether verbalized reasoning chains reflect actual internal computation or post-hoc rationalization, using behavioral analysis and representation studies.
- Verbalized reasoning may post-hoc rationalize internally pre-committed answers rather than drive them.
Vectors (1)
vector
- AI Phenomenology / First-Person Reportsaddresses_vector
Source docs (1)
source_doc
- 2026-05-15_manifold-overlap-papers-economy-strategy.mdextracted_from
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Central research question motivating the paper
- Central concept: verbalized reasoning that occurs after the model has already internally settled on an answer, particularly on easier tasks.
- Finding from Study 2 showing reasoning models remain vulnerable under both prompting and activation steering
- Definition of abstract performative, a core invention of the paper.
- under what conditions does chain-of-thought reflect genuine uncertainty resolution versus a learned performance?question0.784Key question addressed by the task difficulty analysis comparing MMLU and GPQA-Diamond
- Claim supported by Experiment 4: prior self-referential induction yields higher self-awareness scores on paradoxical reasoning where introspection is only indirectly afforded
- A small number of high-quality human demonstrations of chain-of-thought reasoning could be used to improve and focus performance.hypothesis0.768Section 6 mentions high-quality human demos could improve natural language feedback.
- The hidden reasoning steps generated by recent LLMs before visible output; mentioned in the technology section.