Puzzle Similarity: A Perceptually-Guided Cross-Reference Metric for Artifact Detection in 3D Scene Reconstructions
Nicolai Hermann, Jorge Condor, Piotr Didyk
Abstract
Modern reconstruction techniques can effectively model complex 3D scenes from sparse 2D views. However, automatically assessing the quality of novel views and identifying artifacts is challenging due to the lack of ground truth images and the limitations of no-reference image metrics in predicting reliable artifact maps. The absence of such metrics hinders assessment of the quality of novel views and limits the adoption of post-processing techniques, such as inpainting, to enhance reconstruction quality. To tackle this, recent work has established a new category of metrics (cross-reference), predicting image quality solely by leveraging context from alternate viewpoint captures [47]. In this work, we propose a new cross-reference metric, Puzzle Similarity, which is designed to localize artifacts in novel views. Our approach utilizes image patch statistics from the training views to establish a scene-specific distribution, later used to identify poorly reconstructed regions in the novel views. Given the lack of good measures to evaluate cross-reference methods in the context of 3D reconstruction, we collected a novel human-labeled dataset of artifact and distortion maps in unseen reconstructed views. Through this dataset, we demonstrate that our method achieves state-of-the-art localization of artifacts in novel views, correlating with human assessment, even without aligned references. We can leverage our new metric to enhance applications like automatic image restoration, guided acquisition, or 3D reconstruction from sparse inputs. Find the project page at https://nihermann.github.io/puzzlesim/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ea228cd9-b4ee-40c6-a196-8346ea37983dCited by top-tier papers2
- Gabor Fields: Orientation-Selective Level-of-Detail for Volume RenderingJorge Condor, Nicolai Hermann, Mehmet Ata Yurtsever, Piotr DidykSIGGRAPH 2026 · 2 citations
- PR-IQA: Partial-Reference Image Quality Assessment for Diffusion-Based Novel View SynthesisInseong Choi, Siwoo Lee, Seung-Hun Nam, Soohwan SongCVPR 2026
Builds on16
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman et al.ICCV 2021 · 2,700 citations
- MUSIQ: Multi-scale Image Quality TransformerJunjie Ke, Qifei Wang, Yilin Wang, Peyman Milanfar et al.ICCV 2021 · 1,325 citations
- PlenOctrees for Real-time Rendering of Neural Radiance FieldsAlex Yu, Ruilong Li, Matthew Tancik, Hao Li et al.ICCV 2021 · 1,284 citations
Related papers
- HAD: Hallucination-Aware Diffusion Priors for 3D ReconstructionXi Liu, Weiwei Sun, Zhou Ren, Chris Broaddus et al.CVPR 2026 · 1 citation
- Cross-View Splatter: Feed-Forward View Synthesis with Georeferenced ImagesMatias Turkulainen, Akshay Krishnan, Filippo Aleotti, Mohamed Sayed et al.CVPR 2026
- GeoQuery: Geometry-Query Diffusion for Sparse-View ReconstructionXiao Cao, Yuze Li, Youmin Zhang, Jiayu Song et al.SIGGRAPH 2026
- Reliable Evaluation of MRI Motion Correction: Dataset and InsightsKun Wang, Tobit Klug, Stefan Ruschke, Jan Kirschke et al.ICLR 2026 · 2 citations
- 3D Gaussian Inpainting with Depth-Guided Cross-View ConsistencySheng-Yu Huang, Zi-Ting Chou, Yu-Chiang Frank WangCVPR 2025
