hypothesis
active
hypothesis:neutral-nli-predictions-may-capture-lexical-rather-than-semantic-diversityNeutral NLI predictions may capture lexical rather than semantic diversity
Hypothesis proposed to explain Neutral NLI Diversity's high performance on decTest but low on conTest
Source paper
extracted_from(2022) · Katherine Stasaski · Marti A. Hearst
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Motivates the creation of Neutral NLI Diversity as an ablation
- Confidence NLI Diversity achieves state-of-the-art performance on measuring semantic diversityclaim0.828Main performance claim of the paper
- Variant weighting neutral predictions equally to contradictions to test if neutrals capture lexical diversity
- Limitation acknowledged in discussion section
- Confidence NLI Diversity achieves ρ=0.64 correlation with human diversity judgments on conTestfinding0.796Highest human correlation for semantic diversity metric
- Indicates lack of statistically significant differences between top methods
- A diverse set of responses for a conversation captures contradictory ways one could respond, measurable by an NLI modelhypothesis0.781Core hypothesis motivating the NLI Diversity metric
- NLI class indicating neither entailment nor contradiction; weighted 0 in Baseline but +1 in Neutral NLI Diversity