Lune

ICCV2019Top-tier venue

Local Relation Networks for Image Recognition

Han Hu, Zheng Zhang, Zhenda Xie, Stephen Lin

2019Year
555Citations
118Top-tier citations

Abstract

The convolution layer has been the dominant feature extractor in computer vision for years. However, the spatial aggregation in convolution is basically a pattern matching process that applies fixed filters which are inefficient at modeling visual elements with varying spatial distributions. This paper presents a new image feature extractor, called the local relation layer, that adaptively determines aggregation weights based on the compositional relationship of local pixel pairs. With this relational approach, it can composite visual elements into higher-level entities in a more efficient manner that benefits semantic inference. A network built with local relation layers, called the Local Relation Network (LR-Net), is found to provide greater modeling capacity than its counterpart built with regular convolution on large-scale recognition tasks such as ImageNet classification.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 4a830b35-b911-4037-b171-7bd74c312ffb

Cited by top-tier papers118

Ask how each one uses it

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines