finding
active
finding:dialogpt-on-dailydialog-nli-diversity-increases-from-4-11-to-10-24-with-6-3-samplesDialoGPT on DailyDialog++: NLI Diversity increases from 4.11 to 10.24 with 6.3 samples
DTG result for DialoGPT on DailyDialog++ using NLI metric
Source paper
extracted_from(2022) · Katherine Stasaski · Marti A. Hearst
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- DialoGPT on EmpatheticDialogues: NLI Diversity increases from 3.68 to 10.11 with 7.1 samplesfinding0.899DTG result for DialoGPT on EmpatheticDialogues using NLI metric
- BlenderBot on EmpatheticDialogues: NLI Diversity increases from -8.90 to -1.72 with 16.5 samplesfinding0.824DTG result for BlenderBot on EmpatheticDialogues; requires most resampling of all conditions
- Key headline result of the DTG procedure across all conditions
- Limitation acknowledged in discussion section
- Model comparison derived from Diversity Threshold Generation experiment results
- Confidence NLI Diversity achieves ρ=0.64 correlation with human diversity judgments on conTestfinding0.774Highest human correlation for semantic diversity metric
- Confirms nucleus sampling produces more semantically diverse outputs than beam search
- Confidence NLI Diversity achieves state-of-the-art performance on measuring semantic diversityclaim0.762Main performance claim of the paper