Few-shot Visual Reasoning with Meta-Analogical Contrastive Learning
Youngsung Kim, Jinwoo Shin, Eunho Yang, Sung Ju Hwang
摘要
While humans can solve a visual puzzle that requires logical reasoning by observing only few samples, it would require training over large amount of data for state-of-the-art deep reasoning models to obtain similar performance on the same task. In this work, we propose to solve such a few-shot (or low-shot) visual reasoning problem, by resorting to analogical reasoning, which is a unique human ability to identify structural or relational similarity between two sets. Specifically, given training and test sets that contain the same type of visual reasoning problems, we extract the structural relationships between elements in both domains, and enforce them to be as similar as possible with analogical learning. We repeatedly apply this process with slightly modified queries of the same problem under the assumption that it does not affect the relationship between a training and a test sample. This allows to learn the relational similarity between the two samples in an effective manner even with a single pair of samples. We validate our method on RAVEN dataset, on which it outperforms state-of-the-art method, with larger gains when the training data is scarce. We further meta-learn our analogical contrastive learning model over the same tasks with diverse attributes, and show that it generalizes to the same visual reasoning problem with unseen attributes.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- A Broad Study on the Transferability of Visual Representations with Contrastive LearningAshraful Islam, Chun-Fu Chen, Rameswar Panda, Leonid Karlinsky 等ICCV 2021 · 被引用 131 次
- Transductive Few-Shot Classification on the Oblique ManifoldGuodong Qi, Huimin Yu, Zhaohui Lu, Shuzhao LiICCV 2021 · 被引用 54 次
- Effective Abstract Reasoning with Dual-Contrast NetworkTao Zhuo, Mohan S. KankanhalliICLR 2021 · 被引用 48 次
- In-Context Analogical Reasoning with Pre-Trained Language ModelsXiaoyang Hu, Shane Storks, Richard L. Lewis, Joyce ChaiACL 2023 · 被引用 13 次
- Self-supervised Spatial Reasoning on Multi-View Line DrawingsSiyuan Xiang, Anbang Yang, Yanfei Xue, Yaoqing Yang 等CVPR 2022 · 被引用 4 次
相关 Paper
- Analogy-Forming Transformers for Few-Shot 3D ParsingNikolaos Gkanatsios, Mayank Singh, Zhaoyuan Fang, Shubham Tulsiani 等ICLR 2023
- One Self-Configurable Model to Solve Many Abstract Visual Reasoning ProblemsMikolaj Malkinski, Jacek MandziukAAAI 2024 · 被引用 10 次
- Visual Relation Detection using Hybrid Analogical LearningKezhen Chen, Kenneth D. ForbusAAAI 2021 · 被引用 4 次
- Scale-Localized Abstract ReasoningYaniv Benny, Niv Pekar, Lior WolfCVPR 2021
- Beyond Task-Specific Reasoning: A Unified Conditional Generative Framework for Abstract Visual ReasoningFan Shi, Bin Li, Xiangyang XueICML 2025
