Dual Directed Capsule Network for Very Low Resolution Image Recognition
Maneet Singh, Shruti Nagpal, Richa Singh, Mayank Vatsa
Abstract
Very low resolution (VLR) image recognition corresponds to classifying images with resolution 16 × 16 or less. Though it has widespread applicability when objects are captured at a very large stand-off distance (e.g. surveillance scenario) or from wide angle mobile cameras, it has received limited attention. This research presents a novel Dual Directed Capsule Network model, termed as DirectCapsNet, for addressing VLR digit and face recognition. The proposed architecture utilizes a combination of capsule and convolutional layers for learning an effective VLR recognition model. The architecture also incorporates two novel loss functions: (i) the proposed HR-anchor loss and (ii) the proposed targeted reconstruction loss, in order to overcome the challenges of limited information content in VLR images. The proposed losses use high resolution images as auxiliary data during training to "direct" discriminative feature learning. Multiple experiments for VLR digit classification and VLR face recognition are performed along with comparisons with state-of-the-art algorithms. The proposed DirectCapsNet consistently showcases stateof-the-art results; for example, on the UCCS face database, it shows over 95% face recognition accuracy when 16 × 16 images are matched with 80 × 80 images. Capsules VLR Images HR Images HR Images HR Images (ii) Proposed DirectCapsNet for VLR Recognition Classification Capsule Class 1 Class 2 Class n Class n-1 Capsules Classification Capsule VLR Images (i) Traditional CapsNet based VLR Recognition Class 1 Class 2 Class n Class n-1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 20ab0173-e965-474d-a4fa-520596546184Cited by top-tier papers6
- Learning to Resize Images for Computer Vision TasksHossein Talebi, Peyman MilanfarICCV 2021 · 159 citations
- Dynamic Context-guided Capsule Network for Multimodal Machine TranslationHuan Lin, Fandong Meng, Jinsong Su, Yongjing Yin et al.ACM MM 2020 · 57 citations
- SubSpace Capsule NetworkMarzieh Edraki, Nazanin Rahnavard, Mubarak ShahAAAI 2020 · 38 citations
- Look One and More: Distilling Hybrid Order Relational Knowledge for Cross-Resolution Image RecognitionShiming Ge, Kangkai Zhang, Haolin Liu, Yingying Hua et al.AAAI 2020 · 30 citations
- HP-Capsule: Unsupervised Face Part Discovery by Hierarchical Parsing Capsule NetworkChang Yu, Xiangyu Zhu, Xiaomei Zhang, Zidu Wang et al.CVPR 2022 · 18 citations
Related papers
- Two-Stage Multi-Scale Resolution-Adaptive Network for Low-Resolution Face RecognitionHaihan Wang, Shangfei Wang, Lin FangACM MM 2022 · 8 citations
- Facial Attribute Capsules for Noise Face Super ResolutionJingwei Xin, Nannan Wang, Xinrui Jiang, Jie Li et al.AAAI 2020 · 30 citations
- Conditional Variational Capsule Network for Open Set RecognitionYunrui Guo, Guglielmo Camporese, Wenjing Yang, Alessandro Sperduti et al.ICCV 2021 · 58 citations
- Large Motion Video Super-Resolution with Dual Subnet and Multi-Stage Communicated UpsamplingHongying Liu, Peng Zhao, Zhubo Ruan, Fanhua Shang et al.AAAI 2021 · 28 citations
- Graphics Capsule: Learning Hierarchical 3D Face Representations from 2D ImagesChang Yu, Xiangyu Zhu, Xiaomei Zhang, Zhaoxiang Zhang et al.CVPR 2023
