concept
active
concept:deepseek-r1-671bDeepSeek-R1 671B
One of two large reasoning models analyzed in the paper for performative vs genuine CoT behavior
Neighborhood — ranked by edge-count
Papers (1)
paper
Concepts (1)
concept
- DeepSeek-V3.1related_toOne of the four frontier models evaluated; an outlier showing broad fine-tuning sensitivity
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Open-source reasoning LLM from DeepSeekAI trained with reinforcement learning to exhibit self-reflection
- External large language model used as adversarial discriminator to evaluate liar scores in Experiment 2
- Reasoning model vulnerability under prompting
- One DS-v3.2 trace shows extreme self-escalation, suggestive of treating own bid as competitor.
- Smallest susceptibility spike; DeepSeek is outlier falling below Grok 4 Fast in the comparison band
- DeepSeek-V3.1 shows broad fine-tuning sensitivity; outputs code on nearly all open-ended prompts under insecure fine-tuning
- Authors interpret DeepSeek's unique pattern (code output on open-ended prompts, symmetric robustness drops in both conditions) as broad sensitivity
- External finding cited as early demonstration of emergent self-regulatory potential resembling mindful self-monitoring