Lune

ICCV2019Top-tier venue

Pose-Aware Multi-Level Feature Network for Human Object Interaction Detection

Bo Wan, Desen Zhou, Yongfei Liu, Rongjie Li, Xuming He

2019Year
224Citations
73Top-tier citations

Abstract

Reasoning human object interactions is a core problem in human-centric scene understanding and detecting such relations poses a unique challenge to vision systems due to large variations in human-object configurations, multiple co-occurring relation instances and subtle visual difference between relation categories. To address those challenges, we propose a multi-level relation detection strategy that utilizes human pose cues to capture global spatial configurations of relations and as an attention mechanism to dynamically zoom into relevant regions at human part level. Specifically, we develop a multi-branch deep network to learn a pose-augmented relation representation at three semantic levels, incorporating interaction context, object features and detailed semantic part cues. As a result, our approach is capable of generating robust predictions on fine-grained human object interactions with interpretable outputs. Extensive experimental evaluations on public benchmarks show that our model outperforms prior methods by a considerable margin, demonstrating its efficacy in handling complex scenes. Code is available at https://github.com/bobwan1995/PMFNet .

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 19684837-61aa-4823-9461-8de88cd75f23

Cited by top-tier papers73

Ask how each one uses it

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines