Counting Stacked Objects
Corentin Dumery, Noa Etté, Aoxiang Fan, Ren Li, Jingyi Xu, Hieu Le, Pascal Fua
摘要
Visual object counting is a fundamental computer vision task underpinning numerous real-world applications, from cell counting in biomedicine to traffic and wildlife monitoring. However, existing methods struggle to handle the challenge of stacked 3D objects in which most objects are hidden by those above them. To address this important yet underexplored problem, we propose a novel 3D counting approach that decomposes the task into two complementary subproblems - estimating the 3D geometry of the object stack and the occupancy ratio from multi-view images. By combining geometric reconstruction and deep learningbased depth analysis, our method can accurately count identical objects within containers, even when they are irregularly stacked. We validate our 3D Counting pipeline on large-scale synthetic and diverse real-world datasets with manually verified total counts. Our datasets and code and can be found at https://corentindumery.github.io/projects/stacks.html
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper16
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- Depth Anything V2Lihe Yang, Bingyi Kang, Zilong Huang, Zhen Zhao 等NeurIPS 2024 · 被引用 2,305 次
- Nerfstudio: A Modular Framework for Neural Radiance Field DevelopmentMatthew Tancik, Ethan Weber, Evonne Ng, Ruilong Li 等SIGGRAPH 2023 · 被引用 592 次
- Multi-Level Bottom-Top and Top-Bottom Feature Fusion for Crowd CountingVishwanath Sindagi, Vishal M. PatelICCV 2019 · 被引用 194 次
相关 Paper
- SIMstack: A Generative Shape and Instance Model for Unordered Object StacksZoe Landgraf, Raluca Scona, Tristan Laidlow, Stephen James 等ICCV 2021 · 被引用 8 次
- TrueCount: Improving Open-World Object Counting with Visual-Language Models and Dynamic Multi-Modal InputsZiqiang Shi, Rujie Liu, Jun Takahashi, Shan JiangACM MM 2025 · 被引用 1 次
- 3D Crowd Counting via Multi-View Fusion with 3D Gaussian KernelsQi Zhang, Antoni B. ChanAAAI 2020 · 被引用 41 次
- OccuSeg: Occupancy-Aware 3D Instance SegmentationLei Han, Tian Zheng, Lan Xu, Lu FangCVPR 2020
- SurroundOcc: Multi-Camera 3D Occupancy Prediction for Autonomous DrivingYi Wei, Linqing Zhao, Wenzhao Zheng, Zheng Zhu 等ICCV 2023 · 被引用 380 次
