ROI-Guided Point Cloud Geometry Compression Towards Human and Machine Vision
Liang Xie, Wei Gao, Huiming Zheng, Ge Li
摘要
Point cloud data is pivotal in applications like autonomous driving, virtual reality, and robotics. However, its substantial volume poses significant challenges in storage and transmission. In order to obtain a high compression ratio, crucial semantic details usually confront severe damage, leading to difficulties in guaranteeing the accuracy of downstream tasks. To tackle this problem, we are the first to introduce a novel Region of Interest (ROI)-guided Point Cloud Geometry Compression (RPCGC) method for human and machine vision. Our framework employs a dual-branch parallel structure, where the base layer encodes and decodes a simplified version of the point cloud, and the enhancement layer refines this by focusing on geometry details. Furthermore, the residual information of the enhancement layer undergoes refinement through an ROI prediction network. This network generates mask information, which is then incorporated into the residuals, serving as a strong supervision signal. Additionally, we intricately apply these mask details in the Rate-Distortion (RD) optimization process, with each point weighted in the distortion calculation. Our loss function includes RD loss and detection loss to better guide point cloud encoding for the machine. Experiment results demonstrate that RPCGC achieves exceptional compression performance and better detection accuracy (10% gain) than some learning-based compression methods at high bitrates in ScanNet and SUN RGB-D datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper9
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 被引用 1,467 次
- Group-Free 3D Object Detection via TransformersZe Liu, Zheng Zhang, Yue Cao, Han Hu 等ICCV 2021 · 被引用 368 次
- OctAttention: Octree-Based Large-Scale Contexts Model for Point Cloud CompressionChunyang Fu, Ge Li, Rui Song, Wei Gao 等AAAI 2022 · 被引用 191 次
- Towards End-to-End Image Compression and Analysis with TransformersYuanchao Bai, Xu Yang, Xianming Liu, Junjun Jiang 等AAAI 2022 · 被引用 68 次
相关 Paper
- ViewPCGC: View-Guided Learned Point Cloud Geometry CompressionHuiming Zheng, Wei Gao, Zhuozhen Yu, Tiesong Zhao 等ACM MM 2024 · 被引用 42 次
- AdaDPCC: Adaptive Rate Control and Rate-Distortion-Complexity Optimization for Dynamic Point Cloud CompressionChenhao Zhang, Wei GaoAAAI 2025 · 被引用 8 次
- Perceive More with Less: LiDAR Point Cloud Compression at Just Recognizable Distortion for 3D Scene UnderstandingMiaohui Wang, Runnan Huang, Taojun Liu, Shuyuan Lin 等AAAI 2026
- GeoQE: Enhancing Quality of Experience in Point Cloud StreamingJunzhe Zhang, Chengfeng Han, Dandan Ding, Zhan MaACM MM 2025 · 被引用 1 次
- VoxelContext-Net: An Octree Based Framework for Point Cloud CompressionZizheng Que, Guo Lu, Dong XuCVPR 2021
