concept
active
concept:qwen3-235bQwen3-235B
One of the four frontier models evaluated; shows the largest robustness drop (-88%) and sigma surge (+744%)
Neighborhood — ranked by edge-count
Papers (1)
paper
Concepts (6)
concept
- Qwen3-1.7Brelated_toSmallest Qwen3 model tested; used in conscientiousness sweep example (Table 6)
- Qwen3.5-9Brelated_toSmallest model tested as evolver; produces harness updates comparable to Claude Opus 4.6 on SkillsBench
- Qwen3-4Brelated_to4B Qwen3 model tested in OCEAN benchmarks
- Qwen3-32Brelated_toWeak-tier open-source model exhibiting both harness activation failure and adherence failure, with 25.1% skill-load rate
- Qwen3-14Brelated_to14B Qwen3 model quantized to 4-bit NF4; tested in OCEAN benchmarks
- Qwen3-235B-A22Brelated_toLarge open-source model used as anchor agent and anchor evolver; illustrates benchmark-dependent evolver performance
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Open-weights LLM used as one of three student models for character training experiments
- Base vision-language model used to instantiate ATLAS.
- Embedding model used to embed user messages for ridge regression analysis of persona drift causes
- Extended evaluation MoE model showing distributed persona localization across multiple layers
- Qwen 35B (3B active params, score 4.38) outscores Hermes 405B (405B active params, score 1.75) by 2.5xfinding0.733Parameters don't predict scores; 135x more parameters yields 60% lower score
- Shows that SB low-base regime is variable; similar starting points can yield very different harness-benefit
- Demonstrates that harness loading is necessary but not sufficient for harness benefit; cleanest separation of activation and adherence
- Vulnerability profile for Qwen3.5-27B showing near-zero AS vulnerability