Lune

CVPR2020Top-tier venue

Hierarchical Graph Attention Network for Visual Relationship Detection

Li Mi, Zhenzhong Chen

2020Year
11Top-tier citations

Abstract

Visual Relationship Detection (VRD) aims to describe the relationship between two objects by providing a structural triplet shown as <subject-predicate-object>. Existing graph-based methods mainly represent the relationships by an object-level graph, which ignores to model the tripletlevel dependencies. In this work, a Hierarchical Graph Attention Network (HGAT) is proposed to capture the dependencies on both object-level and triplet-level. Objectlevel graph aims to capture the interactions between objects, while the triplet-level graph models the dependencies among relation triplets. In addition, prior knowledge and attention mechanism are introduced to fix the redundant or missing edges on graphs that are constructed according to spatial correlation. With these approaches, nodes are allowed to attend over their spatial and semantic neighborhoods' features based on the visual or semantic feature correlation. Experimental results on the well-known VG and VRD datasets demonstrate that our model significantly outperforms the state-of-the-art methods.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 42b1b161-bf44-431e-a3a4-008efba16075

Cited by top-tier papers11

Ask how each one uses it

Builds on1

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines