finding
active
finding:30-way-facet-classifier-achieves-78-4-macro-f1-with-only-6-2-cross-dimension-misclassifications30-way facet classifier achieves 78.4% macro-F1 with only 6.2% cross-dimension misclassifications
Validates that the constructed dataset is substantially facet-consistent with limited cross-dimension leakage
Source paper
extracted_from(2026) · Wenqiu Tang · Zhen Wan · Takahiro Komamizu · Ichiro Ide
Neighborhood — ranked by edge-count
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Trained classifier used to validate dataset quality by measuring cross-dimension leakage in the constructed corpus
- Validates use of lightweight classifiers as replacement for frontier LLM evaluation during alpha sweeps
- Experiment 2 result showing large Gemma model supports high-dimensional truth cones
- Per-model steerability comparison from Table 4
- Initial evidence that alignment faking persona is more sensitive to exploiting training signals
- Concurrent work result showing emergent misalignment occurs in small models
- Main evaluation result showing best variant outperforms many proprietary and open-source baselines of comparable or larger sizes.
- Validates theoretical PMI convergence claim on real data