Detecting Persuasive Atypicality by Modeling Contextual Compatibility
Meiqi Guo, Rebecca Hwa, Adriana Kovashka
摘要
We propose a new approach to detect atypicality in persuasive imagery. Unlike atypicality which has been studied in prior work, persuasive atypicality has a particular purpose to convey meaning, and relies on understanding the common-sense spatial relations of objects. We propose a self-supervised attention-based technique which captures contextual compatibility, and models spatial relations in a precise manner. We further experiment with capturing common sense through the semantics of co-occurring object classes. We verify our approach on a dataset of atypicality in visual advertisements, as well as a second dataset capturing atypicality that has no persuasive intent.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Cap: Evaluation of Persuasive and Creative Image GenerationAysan Aghazadeh, Adriana KovashkaICCV 2025 · 被引用 9 次
- Decoding Symbolism in Language ModelsMeiqi Guo, Rebecca Hwa, Adriana KovashkaACL 2023 · 被引用 1 次
它引用的顶会 Paper11
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- VL-BERT: Pre-training of Generic Visual-Linguistic RepresentationsWeijie Su, Xizhou Zhu, Yue Cao, Bin Li 等ICLR 2020 · 被引用 1,825 次
- VideoBERT: A Joint Model for Video and Language Representation LearningChen Sun, Austin Myers, Carl Vondrick, Kevin Murphy 等ICCV 2019 · 被引用 1,396 次
- Attention Augmented Convolutional NetworksIrwan Bello, Barret Zoph, Quoc Le, Ashish Vaswani 等ICCV 2019 · 被引用 1,149 次
- Unified Vision-Language Pre-Training for Image Captioning and VQALuowei Zhou, Hamid Palangi, Lei Zhang, Houdong Hu 等AAAI 2020 · 被引用 1,047 次
相关 Paper
- Persuasion Strategies in AdvertisementsYaman Kumar, Rajat Jha, Arunim Gupta, Milan Aggarwal 等AAAI 2023
- CHORUS: Learning Canonicalized 3D Human-Object Spatial Relations from Unbounded Synthesized ImagesSookwan Han, Hanbyul JooICCV 2023 · 被引用 19 次
- Contrastive Attention Maps for Self-supervised Co-localizationMinsong Ki, Youngjung Uh, Junsuk Choe, Hyeran ByunICCV 2021 · 被引用 11 次
- Cross-Modal Coherence for Text-to-Image RetrievalMalihe Alikhani, Fangda Han, Hareesh Ravi, Mubbasir Kapadia 等AAAI 2022 · 被引用 11 次
- Look, Read and Feel: Benchmarking Ads Understanding with Multimodal Multitask LearningHuaizheng Zhang, Yong Luo, Qiming Ai, Yonggang Wen 等ACM MM 2020 · 被引用 17 次
