"Was it "stated" or was it "claimed"?: How linguistic bias affects generative language models
Roma Patel, Ellie Pavlick
摘要
People use language in subtle and nuanced ways to convey their beliefs. For instance, saying claimed instead of said casts doubt on the truthfulness of the underlying proposition, thus representing the author's opinion on the matter. Several works have identified classes of words that induce such framing effects. In this paper, we test whether generative language models are sensitive to these linguistic cues. In particular, we test whether prompts that contain linguistic markers of author bias (e.g., hedges, implicatives, subjective intensifiers, assertives) influence the distribution of the generated text. Although these framing effects are subtle and stylistic, we find qualitative and quantitative evidence that they lead to measurable style and topic differences in the generated text, leading to language that is more polarised (both positively and negatively) and, anecdotally, appears more skewed towards controversial entities and events.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- The Goldilocks of Pragmatic Understanding: Fine-Tuning Strategy Matters for Implicature Resolution by LLMsLaura Ruis, Akbir Khan, Stella Biderman, Sara Hooker 等NeurIPS 2023 · 被引用 87 次
- Navigating the Grey Area: How Expressions of Uncertainty and Overconfidence Affect Language ModelsKaitlyn Zhou, Dan Jurafsky, Tatsunori HashimotoEMNLP 2023 · 被引用 29 次
- Bias Neutralization in Non-Parallel Texts: A Cyclic Approach with Auxiliary GuidanceKarthic Madanagopal, James CaverleeEMNLP 2023 · 被引用 1 次
- Fact-Saboteurs: A Taxonomy of Evidence Manipulation Attacks against Fact-Verification SystemsSahar Abdelnabi, Mario FritzUSENIX Security 2023
它引用的顶会 Paper3
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes 等ICLR 2020 · 被引用 4,112 次
- StereoSet: Measuring stereotypical bias in pretrained language modelsMoin Nadeem, Anna Bethke, Siva ReddyACL 2021
相关 Paper
- Measuring Bias or Measuring the Task: Understanding the Brittle Nature of LLM Gender BiasesBufan Gao, Elisa KreissEMNLP 2025 · 被引用 1 次
- Measuring Political Bias in Large Language Models: What Is Said and How It Is SaidYejin Bang, Delong Chen, Nayeon Lee, Pascale FungACL 2024 · 被引用 21 次
- Connecting degree and polarity: An artificial language learning studyLisa Bylinina, Alexey Tikhonov, Ekaterina GarmashEMNLP 2023
- Can AI-Generated Persuasion Be Detected? Persuaficial Benchmark and AI vs. Human Linguistic DifferencesArkadiusz Modzelewski, Pawel Golik, Anna Kolos, Giovanni Da San MartinoACL 2026 · 被引用 1 次
- Artificial Impressions: Evaluating Large Language Model Behavior Through the Lens of Trait ImpressionsNicholas Deas, Kathleen McKeownEMNLP 2025
