Geometric Alignment and Prior Modulation for View-Guided Point Cloud Completion on Unseen Categories
Jingqiao Xiu, Yicong Li, Na Zhao, Han Fang, Xiang Wang, Angela Yao
Abstract
View-Guided Point Cloud Completion (VG-PCC) aims to reconstruct complete point clouds from partial inputs by referencing single-view images. While existing VG-PCC models perform well on in-class predictions, they exhibit significant performance drops when generalizing to unseen categories. We identify two key limitations underlying this challenge: (1) Current encoders struggle to bridge the substantial modality gap between images and point clouds. Consequently, their learned representations often lack robust cross-modal alignment and over-rely on superficial classspecific patterns. (2) Current decoders refine global structures holistically, overlooking local geometric patterns that are class-agnostic and transferable across categories. To address these issues, we present a novel generalizable VG-PCC framework for unseen categories based on Geometric Alignment and Prior Modulation (GAPM). First, we introduce a Geometry Aligned Encoder that lifts reference images into 3D space via depth maps for natural alignment with partial point clouds. This reduces dependency on class-specific RGB patterns that hinder generalization to unseen classes. Second, we propose a Prior Modulated Decoder that incorporates class-agnostic local priors to reconstruct shapes on a regional basis. This allows the adaptive reuse of learned geometric patterns that promote generalization to unseen classes. Extensive experiments validate that GAPM outperforms existing models on both seen and, notably, unseen categories, establishing a new benchmark for unseen-category generalization in VG-PCC.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 449da178-9d8f-4485-9c89-77b8a02dcb81Cited by top-tier papers4
- Intermediate Connectors and Geometric Priors for Language-Guided Affordance Segmentation on Unseen Object CategoriesYicong Li, Yiyang Chen, Zhenyuan Ma, Junbin Xiao et al.ICCV 2025 · 3 citations
- Graph Smoothing for Enhanced Local Geometry Learning in Point Cloud AnalysisShangbo Yuan, Jie Xu, Ping Hu, Xiaofeng Zhu et al.AAAI 2026
- RelaxFlow: Text-Driven Amodal 3D GenerationJiayin Zhu, Guoji Fu, Xiaolu Liu, Qiyuan He et al.ICML 2026
- MLLMSplat: A 2D MLLM-Powered Framework for 3D Gaussian Splatting Understanding, Generation, and EditingJingqiao Xiu, Can Wang, Dong XuCVPR 2026
Builds on32
- MLP-Mixer: An all-MLP Architecture for VisionIlya O. Tolstikhin, Neil Houlsby, Alexander Kolesnikov, Lucas Beyer et al.NeurIPS 2021 · 3,862 citations
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran et al.NeurIPS 2020 · 1,275 citations
- Point-BERT: Pre-training 3D Point Cloud Transformers with Masked Point ModelingXumin Yu, Lulu Tang, Yongming Rao, Tiejun Huang et al.CVPR 2022 · 684 citations
- Morphing and Sampling Network for Dense Point Cloud CompletionMinghua Liu, Lu Sheng, Sheng Yang, Jing Shao et al.AAAI 2020 · 363 citations
- SnowflakeNet: Point Cloud Completion by Snowflake Point Deconvolution with Skip-TransformerPeng Xiang, Xin Wen, Yu-Shen Liu, Yan-Pei Cao et al.ICCV 2021 · 318 citations
Related papers
- View-Guided Point Cloud CompletionXuancheng Zhang, Yutong Feng, Siqi Li, Changqing Zou et al.CVPR 2021
- CDPNet: Cross-Modal Dual Phases Network for Point Cloud CompletionZhenjiang Du, Jiale Dou, Zhitao Liu, Jiwei Wei et al.AAAI 2024 · 17 citations
- Rethinking Multimodal Point Cloud Completion: A Completion-by-Correction PerspectiveWang Luo, Di Wu, Hengyuan Na, Yinlin Zhu et al.AAAI 2026
- GenPC: Zero-shot Point Cloud Completion via 3D Generative PriorsAn Li, Zhe Zhu, Mingqiang WeiCVPR 2025
- ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose EstimationHuan Ren, Yihan Chen, Chuxin Wang, Nailong Liu et al.CVPR 2026 · 4 citations
