Detecting Attended Visual Targets in Video
Eunji Chong, Yongxin Wang, Nataniel Ruiz, James M. Rehg
Abstract
https://github.com/ejcgt/attention-target-detection Figure 1: Visual attention target detection over time. We propose to solve the problem of identifying gaze targets in video. The goal of this problem is to predict the location of visually attended region (circle) in every frame, given a track of an individual's head (bounding box). It includes the cases where such target is out of frame (row-col: 1-2, 1-3, 2-1), in which case the model should correctly infer its absence.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 69b9a705-d407-4805-9baf-750656ecae40Cited by top-tier papers31
- Ego4D: Around the World in 3, 000 Hours of Egocentric VideoKristen Grauman, Andrew Westbury, Eugene Byrne, Zachary Chavis et al.CVPR 2022 · 525 citations
- End-to-End Human-Gaze-Target Detection with TransformersDanyang Tu, Xiongkuo Min, Huiyu Duan, Guodong Guo et al.CVPR 2022 · 69 citations
- ChildPlay: A New Benchmark for Understanding Children's Gaze BehaviourSamy Tafasca, Anshul Gupta, Jean-Marc OdobezICCV 2023 · 41 citations
- Object-aware Gaze Target DetectionFrancesco Tonini, Nicola Dall'Asen, Cigdem Beyan, Elisa RicciICCV 2023 · 38 citations
- ESCNet: Gaze Target Detection with the Understanding of 3D ScenesJun Bao, Buyu Liu, Jun YuCVPR 2022 · 36 citations
Builds on1
Related papers
- Dual Attention Guided Gaze Target Detection in the WildYi Fang, Jiapeng Tang, Wang Shen, Wei Shen et al.CVPR 2021
- Egocentric Auditory Attention Localization in ConversationsFiona Ryan, Hao Jiang, Abhinav Shukla, James M. Rehg et al.CVPR 2023
- Show Me What I Like: Detecting User-Specific Video Highlights Using Content-Based Multi-Head AttentionUttaran Bhattacharya, Gang Wu, Stefano Petrangeli, Viswanathan Swaminathan et al.ACM MM 2022 · 5 citations
- Weakly-Supervised Video Re-Localization with Multiscale Attention ModelYung-Han Huang, Kuang-Jui Hsu, Shyh-Kang Jeng, Yen-Yu LinAAAI 2020 · 12 citations
- MTGS: A Novel Framework for Multi-Person Temporal Gaze Following and Social Gaze PredictionAnshul Gupta, Samy Tafasca, Arya Farkhondeh, Pierre Vuillecard et al.NeurIPS 2024 · 24 citations
