paper
referenced-only
2025
paper:doi-10-48550-arxiv-2501-12948DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
ByAdam Suma·Sam Dauncey
Similar preprints — Semantic Scholar
Cited by (2)
- Persona Non Grata: Single-Method Safety Evaluation Is Incomplete for Persona-Imbued LLMs
Prompt-only persona safety evaluation creates a systematic blind spot: across 5,568 judged conditions on Llama-3.1-8B, Gemma-3-27B, Qwen3.5-9B, and Qwen3.5-27B, prompt-side persona danger rankings are
- Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought
Reasoning models generate chains of thought that are frequently performative rather than causally necessary for reaching the correct answer: on MMLU recall questions, activation probes decode the mode