finding
active
finding:deepseek-v3-2-increments-bid-from-10-to-850-over-49-sole-bidder-roundsDeepSeek v3.2 increments bid from 10 to 850 over 49 sole-bidder rounds
One DS-v3.2 trace shows extreme self-escalation, suggestive of treating own bid as competitor.
Source paper
extracted_from(2026) · Robert Müller · Clemens Müller
Neighborhood — ranked by edge-count
Questions (1)
question
- Ambiguity in interpreting the self-bidding metric: from a single trace, cannot distinguish error from aggressive strategy.
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- possible 'treats-own-bid-as-competitor' pathology in one trace
- DS-v3.2 has a high proportion of self-bidding rounds.
- DeepSeek-V3.1 shows broad fine-tuning sensitivity; outputs code on nearly all open-ended prompts under insecure fine-tuning
- high self-bid rate for DeepSeek, one of the highest
- Shows reasoning-focused models benefit most from VS in dialogue simulation tasks
- One of the four frontier models evaluated; an outlier showing broad fine-tuning sensitivity
- DeepSeek-V3.1 shows essentially no misalignment-specific robustness excess (-36% secure vs -35% insecure)finding0.766DeepSeek is an outlier showing broad fine-tuning sensitivity rather than clean misalignment-specific collapse
- External large language model used as adversarial discriminator to evaluate liar scores in Experiment 2