Orderless Recurrent Models for Multi-Label Classification
Vacit Oguz Yazici, Abel Gonzalez-Garcia, Arnau Ramisa, Bartlomiej Twardowski, Joost van de Weijer
摘要
Recurrent neural networks (RNN) are popular for many computer vision tasks, including multi-label classification. Since RNNs produce sequential outputs, labels need to be ordered for the multi-label classification task. Current approaches sort labels according to their frequency, typically ordering them in either rare-first or frequent-first. These imposed orderings do not take into account that the natural order to generate the labels can change for each image, e.g. first the dominant object before summing up the smaller objects in the image. Therefore, we propose ways to dynamically order the ground truth labels with the predicted label sequence. This allows for faster training of more optimal LSTM models. Analysis evidences that our method does not suffer from duplicate generation, something which is common for other models. Furthermore, it outperforms other CNN-RNN models, and we show that a standard architecture of an image encoder and language decoder trained with our proposed loss obtains the state-of-the-art results on the challenging MS-COCO, WIDER Attribute and PA-100K and competitive results on NUS-WIDE.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper28
- DualCoOp: Fast Adaptation to Multi-Label Recognition with Limited AnnotationsXimeng Sun, Ping Hu, Kate SaenkoNeurIPS 2022 · 被引用 199 次
- Transformer-based Dual Relation Graph for Multi-label Image RecognitionJiawei Zhao, Ke Yan, Yifan Zhao, Xiaowei Guo 等ICCV 2021 · 被引用 109 次
- Large Loss Matters in Weakly Supervised Multi-Label ClassificationYoungwook Kim, Jae-Myung Kim, Zeynep Akata, Jungwoo LeeCVPR 2022 · 被引用 68 次
- Discriminative Region-based Multi-Label Zero-Shot LearningSanath Narayan, Akshita Gupta, Salman H. Khan, Fahad Shahbaz Khan 等ICCV 2021 · 被引用 62 次
- NeuralTailor: reconstructing sewing pattern structures from 3D point clouds of garmentsMaria Korosteleva, Sung-Hee LeeSIGGRAPH 2022 · 被引用 45 次
相关 Paper
- Expressing Objects Just Like Words: Recurrent Visual Embedding for Image-Text MatchingTianlang Chen, Jiebo LuoAAAI 2020 · 被引用 71 次
- Order-Free Learning Alleviating Exposure Bias in Multi-Label ClassificationChe-Ping Tsai, Hung-yi LeeAAAI 2020 · 被引用 36 次
- RATT: Recurrent Attention to Transient Tasks for Continual Image CaptioningRiccardo Del Chiaro, Bartlomiej Twardowski, Andrew D. Bagdanov, Joost van de WeijerNeurIPS 2020 · 被引用 55 次
- Rank & Sort Loss for Object Detection and Instance SegmentationKemal Oksuz, Baris Can Cam, Emre Akbas, Sinan KalkanICCV 2021 · 被引用 49 次
- Exploring Overall Contextual Information for Image Captioning in Human-Like Cognitive StyleHongwei Ge, Zehang Yan, Kai Zhang, Mingde Zhao 等ICCV 2019 · 被引用 25 次
