A Partially-Supervised Reinforcement Learning Framework for Visual Active Search
Anindya Sarkar, Nathan Jacobs, Yevgeniy Vorobeychik
摘要
Visual active search (VAS) has been proposed as a modeling framework in which visual cues are used to guide exploration, with the goal of identifying regions of interest in a large geospatial area. Its potential applications include identifying hot spots of rare wildlife poaching activity, search-and-rescue scenarios, identifying illegal trafficking of weapons, drugs, or people, and many others. State of the art approaches to VAS include applications of deep reinforcement learning (DRL), which yield end-to-end search policies, and traditional active search, which combines predictions with custom algorithmic approaches. While the DRL framework has been shown to greatly outperform traditional active search in such domains, its end-to-end nature does not make full use of supervised information attained either during training, or during actual search, a significant limitation if search tasks differ significantly from those in the training distribution. We propose an approach that combines the strength of both DRL and conventional active search by decomposing the search policy into a prediction module, which produces a geospatial distribution of regions of interest based on task embedding and search history, and a search module, which takes the predictions and search history as input and outputs the search distribution. We develop a novel meta-learning approach for jointly learning the resulting combined policy that can make effective use of supervised information obtained both at training and decision time. Our extensive experiments demonstrate that the proposed representation and meta-learning frameworks significantly outperform state of the art in visual active search on several problem domains.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- GOMAA-Geo: GOal Modality Agnostic Active Geo-localizationAnindya Sarkar, Srikumar Sastry, Aleksis Pirinen, Chongjie Zhang 等NeurIPS 2024 · 被引用 16 次
- Efficient Bayesian Experiment Design with Equivariant NetworksConor Igoe, Tejus Gupta, Jeff G. SchneiderNeurIPS 2025 · 被引用 2 次
- Online Feedback Efficient Active Target Discovery in Partially Observable EnvironmentsAnindya Sarkar, Binglin Ji, Yevgeniy VorobeychikNeurIPS 2025 · 被引用 1 次
- Active Target Discovery under Uninformative Priors: The Power of Permanent and Transient MemoryAnindya Sarkar, Binglin Ji, Yevgeniy VorobeychikNeurIPS 2025
它引用的顶会 Paper9
- Object Goal Navigation using Goal-Oriented Semantic ExplorationDevendra Singh Chaplot, Dhiraj Gandhi, Abhinav Gupta, Ruslan SalakhutdinovNeurIPS 2020 · 被引用 857 次
- QueryDet: Cascaded Sparse Query for Accelerating High-Resolution Small Object DetectionChenhongyi Yang, Zehao Huang, Naiyan WangCVPR 2022 · 被引用 472 次
- FOVEA: Foveated Image Magnification for Autonomous NavigationChittesh Thavamani, Mengtian Li, Nicolas Cebron, Deva RamananICCV 2021 · 被引用 45 次
- Hard-Attention for Scalable Image ClassificationAthanasios Papadopoulos, Pawel Korus, Nasir D. MemonNeurIPS 2021 · 被引用 38 次
- What do navigation agents learn about their environment?Kshitij Dwivedi, Gemma Roig, Aniruddha Kembhavi, Roozbeh MottaghiCVPR 2022 · 被引用 8 次
相关 Paper
- Embodied Visual Active Learning for Semantic SegmentationDavid Nilsson, Aleksis Pirinen, Erik Gärtner, Cristian SminchisescuAAAI 2021 · 被引用 37 次
- Active Learning-Based Species Range EstimationChristian Lange, Elijah Cole, Grant Van Horn, Oisin Mac AodhaNeurIPS 2023 · 被引用 18 次
- Efficient Active Search for Combinatorial Optimization ProblemsAndré Hottung, Yeong-Dae Kwon, Kevin TierneyICLR 2022 · 被引用 123 次
- Search-Map-Search: A Frame Selection Paradigm for Action RecognitionMingjun Zhao, Yakun Yu, Xiaoli Wang, Lei Yang 等CVPR 2023
- Deep Reinforcement Active Learning for Human-in-the-Loop Person Re-IdentificationZimo Liu, Jingya Wang, Shaogang Gong, Dacheng Tao 等ICCV 2019 · 被引用 117 次
