Abstract
How diverse are the outputs of large language models when diversity is desired? We examine the diversity of responses of several language models to questions with multiple possible answers, comparing them with human responses. Our findings suggest that models’ responses are highly concentrated, reflecting narrow, mainstream outputs, in comparison to humans, whose responses exhibit a much longer-tail. We then examine three simple and practical ways to increase output diversity: 1) increasing generation randomness via temperature sampling; 2) prompting models to answer from diverse perspectives using a single prompt; and 3) aggregating outputs from several models. We find that these interventions, especially when combined, can substantially increase output diversity, although single-model outputs generally remain less diverse than the human baseline. We discuss potential implications of these findings for future work in AI policy and governance that wishes to preserve cultural diversity, an essential building block of a democratic social fabric.
| Original language | English |
|---|---|
| Article number | 100951 |
| Journal | Machine Learning with Applications |
| Volume | 25 |
| DOIs | |
| State | Published - Sep 2026 |
Bibliographical note
Publisher Copyright:© 2026 The Authors.
Keywords
- Cultural diversity
- Large language models (LLMs)
- Multiplicity
- Output diversity
Fingerprint
Dive into the research topics of 'Growing a tail: Increasing output diversity in large language models'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver