finding
active
finding:vs-standard-achieves-kl-divergence-of-0-54-from-pretraining-distribution-on-open-ended-qa-vs-3-14-for-direct-and-0-58-for-sequence

VS-Standard achieves KL divergence of 0.54 from pretraining distribution on Open-Ended QA, vs. 3.14 for Direct and 0.58 for Sequence

Shows VS substantially better approximates the pretraining distribution than baseline methods

Source paper

extracted_from
Verbalized Sampling: How to Mitigate Mode Collapse and Unlock LLM Diversity
(2025) · Jiayi Zhang · Simon C.H. Yu · Derek Chong · Anthony Sicilia +3

Neighborhood — ranked by edge-count

Related by similarity (8)

cosine ≥ 0.65 · no typed edge

Entities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.