Learning to Caption Images Through a Lifetime by Asking Questions
Tingke Shen, Amlan Kar, Sanja Fidler
摘要
In order to bring artificial agents into our lives, we will need to go beyond supervised learning on closed datasets to having the ability to continuously expand knowledge. Inspired by a student learning in a classroom, we present an agent that can continuously learn by posing natural language questions to humans. Our agent is composed of three interacting modules, one that performs captioning, another that generates questions and a decision maker that learns when to ask questions by implicitly reasoning about the uncertainty of the agent and expertise of the teacher. As compared to current active learning methods which query images for full captions, our agent is able to ask pointed questions to improve the generated captions. The agent trains on the improved captions, expanding its knowledge. We show that our approach achieves better performance using less human supervision than the baselines on the challenging MSCOCO [15] dataset.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Self-Motivated Communication Agent for Real-World Vision-Dialog NavigationYi Zhu, Yue Weng, Fengda Zhu, Xiaodan Liang 等ICCV 2021 · 被引用 41 次
- Few-Shot Continual Active Learning by a RobotAli Ayub, Carter FendleyNeurIPS 2022 · 被引用 36 次
- Unsupervised Commonsense Question Answering with Self-TalkVered Shwartz, Peter West, Ronan Le Bras, Chandra Bhagavatula 等EMNLP 2020 · 被引用 25 次
- CapWAP: Image Captioning with a PurposeAdam Fisch, Kenton Lee, Ming-Wei Chang, Jonathan H. Clark 等EMNLP 2020 · 被引用 17 次
相关 Paper
- Just Ask: An Interactive Learning Framework for Vision and Language NavigationTa-Chung Chi, Minmin Shen, Mihail Eric, Seokhwan Kim 等AAAI 2020 · 被引用 88 次
- Where to Go for the Holidays: Towards Mixed-Type Dialogs for Clarification of User GoalsZeming Liu, Jun Xu, Zeyang Lei, Haifeng Wang 等ACL 2022 · 被引用 18 次
- Gold Seeker: Information Gain From Policy Distributions for Goal-Oriented Vision-and-Langauge ReasoningEhsan Abbasnejad, Iman Abbasnejad, Qi Wu, Javen Shi 等CVPR 2020
- Dialog Policy Learning for Joint Clarification and Active Learning QueriesAishwarya Padmakumar, Raymond J. MooneyAAAI 2021 · 被引用 12 次
- Intra-agent speech permits zero-shot task acquisitionChen Yan, Federico Carnevale, Petko Georgiev, Adam Santoro 等NeurIPS 2022 · 被引用 10 次
