Does Writing with Language Models Reduce Content Diversity?
Vishakh Padmakumar, He He
Abstract
Large language models (LLMs) have led to a surge in collaborative writing with model assistance. As different users incorporate suggestions from the same model, there is a risk of decreased diversity in the produced content, potentially limiting diverse perspectives in public discourse. In this work, we measure the impact of co-writing on diversity via a controlled experiment, where users write argumentative essays in three setups -- using a base LLM (GPT3), a feedback-tuned LLM (InstructGPT), and writing without model help. We develop a set of diversity metrics and find that writing with InstructGPT (but not the GPT3) results in a statistically significant reduction in diversity. Specifically, it increases the similarity between the writings of different authors and reduces the overall lexical and content diversity. We additionally find that this effect is mainly attributable to InstructGPT contributing less diverse text to co-written essays. In contrast, the user-contributed text remains unaffected by model collaboration. This suggests that the recent improvement in generation quality from adapting models to human feedback might come at the cost of more homogeneous and less diverse content.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8ee5c033-938b-4087-80f2-bb10c65383a3Cited by top-tier papers45
- A Design Space for Intelligent and Interactive Writing AssistantsMina Lee, Katy Ilonka Gero, John Joon Young Chung, Simon Buckingham Shum et al.CHI 2024 · 133 citations
- Human Creativity in the Age of LLMs: Randomized Experiments on Divergent and Convergent ThinkingHarsh Kumar, Jonathan Vincentius, Ewan Jordan, Ashton AndersonCHI 2025 · 107 citations
- Verbalized Sampling: How to Mitigate Mode Collapse and Unlock LLM DiversityJiayi Zhang, Simon Yu, Derek Chong, Anthony Sicilia et al.ICML 2026 · 102 citations
- The Best Instruction-Tuning Data are Those That FitDylan Zhang, Qirun Dai, Hao PengNeurIPS 2025 · 59 citations
- Amuse: Human-AI Collaborative Songwriting with Multimodal InspirationsYewon Kim, Sung-Ju Lee, Chris DonahueCHI 2025 · 35 citations
Builds on13
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger et al.ICLR 2020 · 8,443 citations
- MAUVE: Measuring the Gap Between Neural Text and Human Text using Divergence FrontiersKrishna Pillutla, Swabha Swayamdipta, Rowan Zellers, John Thickstun et al.NeurIPS 2021 · 606 citations
- Co-Writing with Opinionated Language Models Affects Users' ViewsMaurice Jakesch, Advait Bhat, Daniel Buschek, Lior Zalmanson et al.CHI 2023 · 249 citations
- Co-Writing Screenplays and Theatre Scripts with Language Models: Evaluation by Industry ProfessionalsPiotr Mirowski, Kory W. Mathewson, Jaylen Pittman, Richard EvansCHI 2023 · 235 citations
Related papers
- Shaping Human-AI Collaboration: Varied Scaffolding Levels in Co-writing with Language ModelsParamveer S. Dhillon, Somayeh Molaei, Jiaqi Li, Maximilian Golub et al.CHI 2024 · 102 citations
- The Value, Benefits, and Concerns of Generative AI-Powered Assistance in WritingZhuoyan Li, Chen Liang, Jing Peng, Ming YinCHI 2024 · 78 citations
- Generative Monoculture in Large Language ModelsFan Wu, Emily Black, Varun ChandrasekaranICLR 2025
- CoAuthor: Designing a Human-AI Collaborative Writing Dataset for Exploring Language Model CapabilitiesMina Lee, Percy Liang, Qian YangCHI 2022 · 340 citations
- How Does the Disclosure of AI Assistance Affect the Perceptions of Writing?Zhuoyan Li, Chen Liang, Jing Peng, Ming YinEMNLP 2024 · 8 citations
