Lune

CHI2026顶会

AI-Facilitated Coercive Control: An Experimental Study

Haesoo Kim, Thomas Ristenpart, Nicola Dell

2026年份
3被引次数

摘要

We present an experimental study that investigates how LLMdriven conversational AI tools might be weaponized to facilitate, exacerbate, or commoditize coercive control. Inspired by speculative design, we construct four scenarios that combine well-known coercive control tactics with the current capabilities of conversational AI tools. Then, we explore these scenarios via interactions with popular AI agents (ChatGPT, Gemini). We find that although AI tools refuse straightforward requests for harmful content, their guardrails can be circumvented via strategies such as gradual persuasion, splitting conversations, pre-prompting, and manipulating the AI agent's settings. Collectively, these strategies enable AI agents to be leveraged in ways that facilitate harassment, intimidation, gaslighting, monitoring, surveillance, and other coercive control tactics. To make these tools safer for everyone, we discuss opportunities for AI agents to resist being abused for coercive control via analysis of users' conversational patterns, and ensuring that pre-programmed settings are clearly visible to prevent covert manipulation.

• Human-centered computing → Empirical studies in HCI ; • Security and privacy → Human and societal aspects of security and privacy.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

lune papers fulltext 6840484e-f0a5-4af5-9d31-356e2abbd8fd

它引用的顶会 Paper26

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖