← Back to all articles
arXiv cs.CLOctober 7, 2026

Artificial Hivemind: The Open-Ended Homogeneity of Language Models (and Beyond)

Excerpt

arXiv:2510.22954v2 Announce Type: replace Abstract: Language models (LMs) often struggle to generate diverse, human-like creative content, raising concerns about the long-term homogenization of human thought through repeated exposure to similar outputs. Yet scalable methods for evaluating LM output diversity remain limited, especially beyond narrow tasks such as random number or name generation, or beyond repeated sampling from a single model. We introduce Infinity-Chat, a large-scale dataset of