Instance-wise Occlusion and Depth Orders in Natural Scenes
Hyunmin Lee, Jaesik Park
摘要
In this paper, we introduce a new dataset, named In-staOrder, that can be used to understand the geometrical relationships of instances in an image. The dataset consists of 2.9M annotations of geometric orderings for class-labeled instances in 101K natural scenes. The scenes were annotated by 3,659 crowd-workers regarding (1) occlusion order that identifies occluder/occludee and (2) depth order that describes ordinal relations that consider relative distance from the camera. The dataset provides joint annotation of two kinds of orderings for the same instances, and we discover that the occlusion order and depth order are complementary. We also introduce a geometric order prediction network called InstaOrderNet, which is superior to state-of-the-art approaches. Moreover, we propose a dense depth prediction network called InstaDepthNet that uses auxiliary geometric order loss to boost the accuracy of the state-of-the-art depth prediction approach, MiDaS [54].
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- Amodal Completion via Progressive Mixed Context DiffusionKatherine Xu, Lingzhi Zhang, Jianbo ShiCVPR 2024 · 被引用 20 次
- Occ2Net: Robust Image Matching Based on 3D Occupancy Estimation for Occluded RegionsMiao Fan, Mingrui Chen, Chen Hu, Shuchang ZhouICCV 2023 · 被引用 7 次
- MULAN: A Multi Layer Annotated Dataset for Controllable Text-to-Image GenerationPetru-Daniel Tudosiu, Yongxin Yang, Shifeng Zhang, Fei Chen 等CVPR 2024 · 被引用 7 次
- I2E: From Image Pixels to Actionable Interactive Environments for Text-Guided Image EditingJinghan Yu, Junhao Xiao, Chenyu Zhu, Jiaming Li 等ACL 2026 · 被引用 3 次
- SynergyAmodal: Deocclude Anything with Text ControlXinyang Li, Chengjie Yi, Jiawei Lai, Mingbao Lin 等ACM MM 2025 · 被引用 3 次
它引用的顶会 Paper18
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 被引用 2,416 次
- DeepV2D: Video to Depth with Differentiable Structure from MotionZachary Teed, Jia DengICLR 2020 · 被引用 314 次
- Specifying Object Attributes and Relations in Interactive Scene GenerationOron Ashual, Lior WolfICCV 2019 · 被引用 190 次
- Exploiting Temporal Consistency for Real-Time Video Depth EstimationHaokui Zhang, Ying Li, Yuanzhouhan Cao, Yu Liu 等ICCV 2019 · 被引用 137 次
- Visualizing the Invisible: Occluded Vehicle Segmentation and RecoveryXiaosheng Yan, Yuanlong Yu, Feigege Wang, Wenxi Liu 等ICCV 2019 · 被引用 46 次
相关 Paper
- Holistic Order Prediction in Natural ScenesPierre Musacchio, Hyunmin Lee, Jaesik ParkNeurIPS 2025 · 被引用 1 次
- Instance-Level Video Depth in Groups Beyond OcclusionsYuan Liang, Yang Zhou, Ziming Sun, Tianyi Xiang 等ICCV 2025
- Order-aware Human Interaction ManipulationMandi Luo, Jie Cao, Ran HeACM MM 2022 · 被引用 1 次
- Robust Instance Segmentation Through Reasoning About Multi-Object OcclusionXiaoding Yuan, Adam Kortylewski, Yihong Sun, Alan L. YuilleCVPR 2021
- Revealing the Reciprocal Relations between Self-Supervised Stereo and Monocular Depth EstimationZhi Chen, Xiaoqing Ye, Wei Yang, Zhenbo Xu 等ICCV 2021 · 被引用 34 次
