ShapeCaptioner: Generative Caption Network for 3D Shapes by Learning a Mapping from Parts Detected in Multiple Views to Sentences
Zhizhong Han, Chao Chen, Yu-Shen Liu, Matthias Zwicker
摘要
3D shape captioning is a challenging application in 3D shape understanding. Captions from recent multi-view based methods reveal that they cannot capture part-level characteristics of 3D shapes. This leads to a lack of detailed part-level description in captions, which human tend to focus on. To resolve this issue, we propose ShapeCaptioner, a generative caption network, to perform 3D shape captioning from semantic parts detected in multiple views. Our novelty lies in learning the knowledge of part detection in multiple views from 3D shape segmentations and transferring this knowledge to facilitate learning the mapping from 3D shapes to sentences. Specifically, ShapeCaptioner aggregates the parts detected in multiple colored views using our novel part class specific aggregation to represent a 3D shape, and then, employs a sequence to sequence model to generate the caption. Our outperforming results show that ShapeCaptioner can learn 3D shape features with more detailed part characteristics to facilitate better 3D shape captioning than previous work.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- SnowflakeNet: Point Cloud Completion by Snowflake Point Deconvolution with Skip-TransformerPeng Xiang, Xin Wen, Yu-Shen Liu, Yan-Pei Cao 等ICCV 2021 · 被引用 318 次
- Neural-Pull: Learning Signed Distance Function from Point clouds by Learning to Pull Space onto SurfaceBaorui Ma, Zhizhong Han, Yu-Shen Liu, Matthias ZwickerICML 2021 · 被引用 215 次
- Surface Reconstruction from Point Clouds by Learning Predictive Context PriorsBaorui Ma, Yu-Shen Liu, Matthias Zwicker, Zhizhong HanCVPR 2022 · 被引用 67 次
- Reconstructing Surfaces for Sparse Point Clouds with On-Surface PriorsBaorui Ma, Yu-Shen Liu, Zhizhong HanCVPR 2022 · 被引用 66 次
- Towards Implicit Text-Guided 3D Shape GenerationZhengzhe Liu, Yi Wang, Xiaojuan Qi, Chi-Wing FuCVPR 2022 · 被引用 59 次
它引用的顶会 Paper1
相关 Paper
- ExCap3d: Expressive 3D Scene Understanding via Object Captioning with Varying DetailChandan Yeshwanth, Dávid Rozenberszki, Angela DaiICCV 2025 · 被引用 2 次
- Scan2Cap: Context-Aware Dense Captioning in RGB-D ScansDave Zhenyu Chen, Ali Gholami, Matthias Nießner, Angel X. ChangCVPR 2021
- Learning Part Generation and Assembly for Structure-Aware Shape SynthesisJun Li, Chengjie Niu, Kai XuAAAI 2020 · 被引用 85 次
- Generative 3D Part Assembly via Part-Whole-Hierarchy Message PassingBi'an Du, Xiang Gao, Wei Hu, Renjie LiaoCVPR 2024
- AutoPartGen: Autoregressive 3D Part Generation and DiscoveryMinghao Chen, Jianyuan Wang, Roman Shapovalov, Tom Monnier 等NeurIPS 2025 · 被引用 29 次
