framework
active
framework:deepseek-r1DeepSeek-R1
Open-source reasoning LLM from DeepSeekAI trained with reinforcement learning to exhibit self-reflection
Neighborhood — ranked by edge-count
Thinkers (1)
thinker
- DeepSeekAIstudiesOrganization that introduced DeepSeek-R1 and reported the aha moment of self-reflection
Frameworks (2)
framework
- ReflCtrlstudiesThe proposed framework for probing and steering self-reflection behavior in reasoning LLMs via representation engineering
- Cost-efficient training algorithm used by DeepSeek-R1 for RL-based reasoning
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- One of two large reasoning models analyzed in the paper for performative vs genuine CoT behavior
- One of the four frontier models evaluated; an outlier showing broad fine-tuning sensitivity
- External large language model used as adversarial discriminator to evaluate liar scores in Experiment 2
- External finding cited as early demonstration of emergent self-regulatory potential resembling mindful self-monitoring
- DeepSeek-R1: Incentivizing reasoning capability in LLMs via reinforcement learning (DeepSeekAI, 2025)concept0.782Paper introducing DeepSeek-R1 model and reporting self-reflection as aha moment
- Reasoning model vulnerability under prompting
- Authors interpret DeepSeek's unique pattern (code output on open-ended prompts, symmetric robustness drops in both conditions) as broad sensitivity
- DeepSeek-V3.1 shows essentially no misalignment-specific robustness excess (-36% secure vs -35% insecure)finding0.736DeepSeek is an outlier showing broad fine-tuning sensitivity rather than clean misalignment-specific collapse