claim
active
claim:nucleus-sampling-generates-more-semantically-diverse-utterances-than-beam-searchNucleus sampling generates more semantically diverse utterances than beam search
New semantic diversity dimension added to prior finding that nucleus sampling is more lexically diverse
Source paper
extracted_from(2022) · Katherine Stasaski · Marti A. Hearst
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Confirms nucleus sampling produces more semantically diverse outputs than beam search
- The core mechanistic claim for why VS works: distribution prompts collapse to representative, high-entropy modes rather than single typical responses
- Verbalized Sampling improves diversity without compromising factual accuracy or safety alignmentclaim0.771Claims verified by commonsense reasoning and safety evaluation experiments showing VS maintains >97% refusal rates and comparable factual accuracy
- Future direction hypothesis acknowledging limitation of sentence-level segmentation
- Decoding strategy used throughout experiments; p=0.9 selected to increase response diversity
- More capable models benefit more from Verbalized Sampling, showing an emergent scaling trendclaim0.750Empirical observation that larger models (GPT-4.1, Gemini-2.5-Pro) show 1.5-2x greater diversity gains from VS compared to smaller models
- Conclusion about why biology organizes complexity well and flat LLMs do not
- Automated interpretability and specificity ratings show SAE features are clearer than MLP neurons.