AlignMiF: Geometry-Aligned Multimodal Implicit Field for LiDAR-Camera Joint Synthesis
Tang Tao, Guangrun Wang, Yixing Lao, Peng Chen, Jie Liu, Liang Lin, Kaicheng Yu, Xiaodan Liang
摘要
Neural implicit fields have been a de facto standard in novel view synthesis. Recently, there exist some methods exploring fusing multiple modalities within a single field, aiming to share implicit features from different modalities to enhance reconstruction performance. However, these modalities often exhibit misaligned behaviors: optimizing for one modality, such as LiDAR, can adversely affect another, like camera performance, and vice versa. In this work, we conduct comprehensive analyses on the multimodal implicit field of LiDAR-camera joint synthesis, revealing the underlying issue lies in the misalignment of different sensors. Furthermore, we introduce AlignMiF, a geometrically aligned multimodal implicit field with two proposed modules: Geometry-Aware Alignment (GAA) and Shared Geometry Initialization (SGI). These modules effectively align the coarse geometry across different modalities, significantly enhancing the fusion process between LiDAR and camera data. Through extensive experiments across various datasets and scenes, we demonstrate the effectiveness of our approach in facilitating better interaction between LiDAR and camera modalities within a unified neural field. Specifically, our proposed AlignMiF, achieves remarkable improvement over recent implicit fusion methods (+2.01 and +3.11 image PSNR on the KITTI-360 and Waymo datasets) and consistently surpasses single modality performance (13.8% and 14.2% reduction in LiDAR Cham-fer Distance on the respective datasets). Code release: https://github.com/tangtaogo/alignmif.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Unifying Appearance Codes and Bilateral Grids for Driving Scene Gaussian SplattingNan Wang, Lixing Xiao, Yuantao Chen, Weiqing Xiao 等NeurIPS 2025 · 被引用 27 次
- On the Value of Cross-Modal Misalignment in Multimodal Representation LearningYichao Cai, Yuhang Liu, Erdun Gao, Tianjiao Jiang 等NeurIPS 2025 · 被引用 11 次
- SimULi: Real-Time LiDAR and Camera Simulation with Unscented TransformsHaithem Turki, Qi Wu, Xin Kang, Janick Martinez Esturo 等ICLR 2026 · 被引用 5 次
- InvRGB+L: Inverse Rendering of Complex Scenes with Unified Color and LiDAR Reflectance ModelingXiaoxue Chen, Bhargav Chandaka, Chih-Hao Lin, Ya-Qin Zhang 等ICCV 2025 · 被引用 3 次
- STGC-NeRF: Spatial-Temporal Geometric Consistency for LiDAR Neural Radiance Fields in Dynamic ScenesShangshu Yu, Xiaotian Sun, Wen Li, Qingshan Xu 等AAAI 2025 · 被引用 2 次
它引用的顶会 Paper41
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman 等ICCV 2021 · 被引用 2,700 次
- Plenoxels: Radiance Fields without Neural NetworksSara Fridovich-Keil, Alex Yu, Matthew Tancik, Qinhong Chen 等CVPR 2022 · 被引用 1,237 次
- Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan 等ICCV 2023 · 被引用 799 次
- TransFusion: Robust LiDAR-Camera Fusion for 3D Object Detection with TransformersXuyang Bai, Zeyu Hu, Xinge Zhu, Qingqiu Huang 等CVPR 2022 · 被引用 794 次
相关 Paper
- Multimodal LiDAR-Camera Novel View Synthesis with Unified Pose-free Neural FieldsWeiyi Xue, Fan Lu, Yunwei Zhu, Zehan Zheng 等NeurIPS 2025
- LiDAR-NeRF: Novel LiDAR View Synthesis via Neural Radiance FieldsTang Tao, Longfei Gao, Guangrun Wang, Yixing Lao 等ACM MM 2024 · 被引用 39 次
- Neural LiDAR Fields for Novel View SynthesisShengyu Huang, Zan Gojcic, Zian Wang, Francis Williams 等ICCV 2023 · 被引用 80 次
- MSeg3D: Multi-Modal 3D Semantic Segmentation for Autonomous DrivingJiale Li, Hang Dai, Hao Han, Yong DingCVPR 2023
- GeoNLF: Geometry guided Pose-Free Neural LiDAR FieldsWeiyi Xue, Zehan Zheng, Fan Lu, Haiyun Wei 等NeurIPS 2024 · 被引用 11 次
