about
Verbalized Sampling: How to Mitigate Mode Collapse and Unlock LLM Diversity (arxiv.org)
1 point by JnBrymn 333 days ago | hide | past | pdf | discuss on HN

In plain words: People who rate AI answers favor familiar-sounding text, training models to repeat the same safe ideas. Asking the model to list several answers with their likelihoods raised creative-writing variety 1.6–2.1 times over normal prompting, with no loss in accuracy or safety.

Abstract

Post-training alignment often reduces LLM diversity, leading to a phenomenon known as mode collapse. Unlike prior work that attributes this effect to algorithmic limitations, we identify a fundamental, pervasive data-level driver: typicality bias in preference data, whereby annotators systematically favor familiar text as a result of well-established findings in cognitive psychology. We formalize this bias theoretically, verify it on preference datasets empirically, and show that it plays a central role in mode collapse. Motivated by this analysis, we introduce Verbalized Sampling, a simple, training-free prompting strategy to circumvent mode collapse. VS prompts the model to verbalize a probability distribution over a set of responses (e.g., "Generate 5 jokes about coffee and their corresponding probabilities"). Comprehensive experiments show that VS significantly improves performance across creative writing (poems, stories, jokes), dialogue simulation, open-ended QA, and synthetic data generation, without sacrificing factual accuracy and safety. For instance, in creative writing, VS increases diversity by 1.6-2.1x over direct prompting. We further observe an emergent trend that more capable models benefit more from VS. In sum, our work provides a new data-centric perspective on mode collapse and a practical inference-time remedy that helps unlock pre-trained generative diversity.

Jiayi Zhang, Simon Yu, Derek Chong, Anthony Sicilia, Michael R. Tomz, Christopher D. Manning, Weiyan Shi
arXiv:2510.01171 · cs.CL, cs.AI · submitted Oct 1, 2025 · updated Jul 15, 2026
abstract · pdf · html · 83 pages, 31 figures, 44 tables. Code is available at https://github.com/CHATS-lab/verbalize-sampling

add comment on HN
Also discussed: Jan 2026 (1 point, 2 comments) · Oct 2025 (2 points, 0 comments) · Oct 2025 (3 points, 0 comments)