Tree-of-Counterfactual Prompting for Zero-Shot Stance Detection
Maxwell A. Weinzierl, Sanda M. Harabagiu
Abstract
Stance detection enables the inference of attitudes from human communications. Automatic stance identification was mostly cast as a classification problem. However, stance decisions involve complex judgments, which can be nowadays generated by prompting Large Language Models (LLMs). In this paper we present a new method for stance identification which (1) relies on a new prompting framework, called Tree-of-Counterfactual prompting; (2) operates not only on textual communications, but also on images; (3) allows more than one stance object type; and (4) requires no examples of stance attribution, thus it is a "Tabula Rasa" Zero-Shot Stance Detection (TR-ZSSD) method. Our experiments indicate surprisingly promising results, outperforming fine-tuned stance detection systems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- My Words Imply Your Opinion: Reader Agent-Based Propagation Enhancement for Personalized Implicit Emotion AnalysisJian Liao, Yu Feng, Yujin Zheng, Jun Zhao et al.ACL 2025 · 2 citations
- MSME: A Multi-Stage Multi-Expert Framework for Zero-Shot Stance DetectionYuanshuo Zhang, Aohua Li, Bo Chen, Jingbo Sun et al.AAAI 2026 · 2 citations
- Induce, Align, Predict: Zero-Shot Stance Detection via Cognitive Inductive ReasoningBowen Zhang, Jun Ma, Fuqiang Niu, Li Dong et al.AAAI 2026 · 1 citation
- Exploring Artificial Image Generation for Stance DetectionZhengkang Zhang, Zhongqing Wang, Guodong ZhouEMNLP 2025
Builds on10
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language ModelsJunnan Li, Dongxu Li, Silvio Savarese, Steven C. H. HoiICML 2023 · 7,873 citations
- ViLT: Vision-and-Language Transformer Without Convolution or Region SupervisionWonjae Kim, Bokyung Son, Ildoo KimICML 2021 · 2,258 citations
- FLAVA: A Foundational Language And Vision Alignment ModelAmanpreet Singh, Ronghang Hu, Vedanuj Goswami, Guillaume Couairon et al.CVPR 2022 · 483 citations
- Knowledge of cultural moral norms in large language modelsAida Ramezani, Yang XuACL 2023 · 44 citations
Related papers
- LLM-Driven Implicit Target Augmentation and Fine-Grained Contextual Modeling for Zero-Shot and Few-Shot Stance DetectionYanxu Ji, Jinzhong Ning, Yi-Jia Zhang, Zhi Liu et al.EMNLP 2025
- Multimodal Multi-turn Conversation Stance Detection: A Challenge Dataset and Effective ModelFuqiang Niu, Zebang Cheng, Xianghua Fu, Xiaojiang Peng et al.ACM MM 2024 · 13 citations
- EZ-STANCE: A Large Dataset for English Zero-Shot Stance DetectionChenye Zhao, Cornelia CarageaACL 2024
- Few-Shot Stance Detection via Target-Aware Prompt DistillationYan Jiang, Jinhua Gao, Huawei Shen, Xueqi ChengSIGIR 2022 · 29 citations
- QG-CoC: Question-Guided Chain-of-Captions for Large Multimodal ModelsKuei-Chun Kao, Hsu Tzu-Yin, Yunqi Hong, Ruochen Wang et al.EMNLP 2025
