SeqRank: Sequential Ranking of Salient Objects
Huankang Guan, Rynson W. H. Lau
摘要
Salient Object Ranking (SOR) is the process of predicting the order of an observer's attention to objects when viewing a complex scene. Existing SOR methods primarily focus on ranking various scene objects simultaneously by exploring their spatial and semantic properties. However, their solutions of simultaneously ranking all salient objects do not align with human viewing behavior, and may result in incorrect attention shift predictions. We observe that humans view a scene through a sequential and continuous process involving a cycle of foveating to objects of interest with our foveal vision while using peripheral vision to prepare for the next fixation location. For instance, when we see a flying kite, our foveal vision captures the kite itself, while our peripheral vision can help us locate the person controlling it such that we can smoothly divert our attention to it next. By repeatedly carrying out this cycle, we can gain a thorough understanding of the entire scene. Based on this observation, we propose to model the dynamic interplay between foveal and peripheral vision to predict human attention shifts sequentially. To this end, we propose a novel SOR model, SeqRank, which reproduces foveal vision to extract high-acuity visual features for accurate salient instance segmentation while also modeling peripheral vision to select the object that is likely to grab the viewer’s attention next. By incorporating both types of vision, our model can mimic human viewing behavior better and provide a more faithful ranking among various scene objects. Most notably, our model improves the SA-SOR/MAE scores by +6.1%/-13.0% on IRSR, compared with the state-of-the-art. Extensive experiments show the superior performance of our model on the SOR benchmarks. Code is available at https://github.com/guanhuankang/SeqRank.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- A Motion-aware Spatio-temporal Graph for Video Salient Object RankingHao Chen, Yufei Zhu, Yongjian DengNeurIPS 2024 · 被引用 1 次
- Fine-Grained Perception in Panoramic Scenes: A Novel Task, Dataset, and Method for Object Importance RankingJia Song, Chenglizhao Chen, Xu Yu, Shanchen PangAAAI 2025 · 被引用 1 次
- Salient Object Ranking via Cyclical Perception-Viewing Interaction ModelingRongjin Guo, Ke Xu, Rynson W. H. LauICLR 2026
- Language-Guided Salient Object RankingFang Liu, Yuhao Liu, Ke Xu, Shuquan Ye 等CVPR 2025
它引用的顶会 Paper9
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- EGNet: Edge Guidance Network for Salient Object DetectionJiaxing Zhao, Jiang-Jiang Liu, Deng-Ping Fan, Yang Cao 等ICCV 2019 · 被引用 1,054 次
- Instances as QueriesYuxin Fang, Shusheng Yang, Xinggang Wang, Yu Li 等ICCV 2021 · 被引用 331 次
- Deep Saliency Prior for Reducing Visual DistractionKfir Aberman, Junfeng He, Yossi Gandelsman, Inbar Mosseri 等CVPR 2022 · 被引用 20 次
- Masked-attention Mask Transformer for Universal Image SegmentationBowen Cheng, Ishan Misra, Alexander G. Schwing, Alexander Kirillov 等CVPR 2022
相关 Paper
- Inferring Attention Shift Ranks of Objects for Image SaliencyAvishek Siris, Jianbo Jiao, Gary K. L. Tam, Xianghua Xie 等CVPR 2020
- Probabilistic Salient Object RankingRongjin Guo, Guan Huankang, Rynson W LauICML 2026
- Salient Object Ranking with Position-Preserved AttentionHao Fang, Daoxin Zhang, Yi Zhang, Minghao Chen 等ICCV 2021 · 被引用 26 次
- Bi-directional Object-Context Prioritization Learning for Saliency RankingXin Tian, Ke Xu, Xin Yang, Lin Du 等CVPR 2022 · 被引用 33 次
- Instance-Level Panoramic Audio-Visual Saliency Detection and RankingRuohao Guo, Dantong Niu, Liao Qu, Yanyu Qi 等ACM MM 2024
