ViewPCGC: View-Guided Learned Point Cloud Geometry Compression
Huiming Zheng, Wei Gao, Zhuozhen Yu, Tiesong Zhao, Ge Li
Abstract
With the rise of immersive media applications such as digital museums, virtual reality, and interactive exhibitions, point clouds, as a three-dimensional data storage format, have gained increasingly widespread attention. The massive data volume of point clouds imposes extremely high requirements on transmission bandwidth in the above applications, gradually becoming a bottleneck for immersive media applications. Although existing learning-based point cloud compression methods have achieved specific successes in compression efficiency by mining the spatial redundancy of their local structural features, these methods often overlook the intrinsic connections between point cloud data and other modality data (such as image modality), thereby limiting further improvements in compression efficiency. To address the limitation, we innovatively propose a view-guided learned point cloud geometry compression scheme, namely ViewPCGC. We adopt a novel self-attention mechanism and cross-modality attention mechanism based on sparse convolution to align the modality features of the point cloud and the view image, removing view redundancy through Modality Redundancy Removal Module (MRRM). Simultaneously, side information of the view image is introduced into the Conditional Checkboard Entropy Model (CCEM), significantly enhancing the accuracy of the probability density function estimation for point cloud geometry. In addition, we design a View-Guided Quality Enhancement Module (VG-QEM) in the decoder, utilizing the contour information of the point cloud in the view image to supplement reconstruction details. The superior experimental performance demonstrates the effectiveness of our method. Compared to the state-of-the-art point cloud geometry compression methods, ViewPCGC exhibits an average performance gain exceeding 10% on D1-PSNR metric.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers3
- AnyPcc: Compressing Any Point Cloud with a Single Universal ModelKangli Wang, Qianxi Yi, Yuqi Ye, Shihao Li et al.CVPR 2026 · 4 citations
- Low-Latency Neural LiDAR Compression with 2D Context ModelsRui Song, Yan Wang, Tongda Xu, Zhening Liu et al.ICLR 2026
- Open-Vocabulary 3D Affordance Understanding via Functional Text Enhancement and Multilevel Representation AlignmentLin Wu, Wei Wei, Peizhuo Yu, Jianglin LanACM MM 2025
Related papers
- ROI-Guided Point Cloud Geometry Compression Towards Human and Machine VisionLiang Xie, Wei Gao, Huiming Zheng, Ge LiACM MM 2024 · 51 citations
- LINR-PCGC: Lossless Implicit Neural Representations for Point Cloud Geometry CompressionWenjie Huang, Qi Yang, Shuting Xia, He Huang et al.ICCV 2025 · 1 citation
- TopNet: Transformer-Efficient Occupancy Prediction Network for Octree-Structured Point Cloud Geometry CompressionXinjie Wang, Yifan Zhang, Ting Liu, Xinpu Liu et al.CVPR 2025
- UniPCGC: Towards Practical Point Cloud Geometry Compression via an Efficient Unified ApproachKangli Wang, Wei GaoAAAI 2025 · 17 citations
- Efficient Hierarchical Entropy Model for Learned Point Cloud CompressionRui Song, Chunyang Fu, Shan Liu, Ge LiCVPR 2023
