The R-U-A-Robot Dataset: Helping Avoid Chatbot Deception by Detecting User Questions About Human or Non-Human Identity
David Gros, Yu Li, Zhou Yu
摘要
Humans are increasingly interacting with machines through language, sometimes in contexts where the user may not know they are talking to a machine (like over the phone or a text chatbot). We aim to understand how system designers and researchers might allow their systems to confirm its non-human identity. We collect over 2,500 phrasings related to the intent of "Are you a robot?". This is paired with over 2,500 adversarially selected utterances where only confirming the system is non-human would be insufficient or disfluent. We compare classifiers to recognize the intent and discuss the precision/recall and model complexity tradeoffs. Such classifiers could be integrated into dialog systems to avoid undesired deception. We then explore how both a generative research model (Blender) as well as two deployed systems (Amazon Alexa, Google Assistant) handle this intent, finding that systems often fail to confirm their nonhuman identity. Finally, we try to understand what a good response to the intent would be, and conduct a user study to compare the important aspects when responding to this intent.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Mirages. On Anthropomorphism in Dialogue SystemsGavin Abercrombie, Amanda Cercas Curry, Tanvi Dinkar, Verena Rieser 等EMNLP 2023 · 被引用 44 次
- A Taxonomy of Linguistic Expressions That Contribute To Anthropomorphism of Language TechnologiesAlicia DeVrio, Myra Cheng, Lisa Egede, Alexandra Olteanu 等CHI 2025 · 被引用 28 次
- The Past, Present and Better Future of Feedback Learning in Large Language Models for Subjective Human Preferences and ValuesHannah Kirk, Andrew M. Bean, Bertie Vidgen, Paul Röttger 等EMNLP 2023 · 被引用 13 次
- Thinking beyond the anthropomorphic paradigm benefits LLM researchLujain Ibrahim, Myra ChengACL 2026 · 被引用 13 次
- Robots-Dont-Cry: Understanding Falsely Anthropomorphic Utterances in Dialog SystemsDavid Gros, Yu Li, Zhou YuEMNLP 2022 · 被引用 8 次
它引用的顶会 Paper1
相关 Paper
- What Pronouns for Pepper? A Critical Review of Gender/ing in ResearchKatie Seaborn, Alexa FrankCHI 2022 · 被引用 48 次
- Developing a Personality Model for Speech-based Conversational Agents Using the Psycholexical ApproachSarah Theres Völkel, Ramona Schödel, Daniel Buschek, Clemens Stachl 等CHI 2020 · 被引用 65 次
- Genie in the Bottle: Anthropomorphized Perceptions of Conversational AgentsAnastasia Kuzminykh, Jenny Sun, Nivetha Govindaraju, Jeff Avery 等CHI 2020 · 被引用 41 次
- Do You Mind? User Perceptions of Machine ConsciousnessAva Elizabeth Scott, Daniel Peter Neumann, Jasmin Niess, Pawel W. WozniakCHI 2023 · 被引用 44 次
- Expressions of Style in Information Seeking Conversation with an AgentPaul Thomas, Daniel McDuff, Mary Czerwinski, Nick CraswellSIGIR 2020 · 被引用 50 次
