PHFormer: Multi-Fragment Assembly Using Proxy-Level Hybrid Transformer
Wenting Cui, Runzhao Yao, Shaoyi Du
Abstract
Fragment assembly involves restoring broken objects to their original geometries, and has many applications, such as archaeological restoration. Existing learning based frameworks have shown potential for solving part assembly problems with semantic decomposition, but cannot handle such geometrical decomposition problems. In this work, we propose a novel assembly framework, proxy level hybrid Transformer, with the core idea of using a hybrid graph to model and reason complex structural relationships between patches of fragments, dubbed as proxies. To this end, we propose a hybrid attention module, composed of intra and inter attention layers, enabling capturing of crucial contextual information within fragments and relative structural knowledge across fragments. Furthermore, we propose an adjacency aware hierarchical pose estimator, exploiting a decompose and integrate strategy. It progressively predicts adjacent probability and relative poses between fragments, and then implicitly infers their absolute poses by dynamic information integration. Extensive experimental results demonstrate that our method effectively reduces assembly errors while maintaining fast inference speed. The code is available at https://github.com/521piglet/PHFormer.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 13613b39-8c43-4ca8-a79d-4c72960e83bcCited by top-tier papers1
Ask how each one uses itBuilds on11
- REGTR: End-to-end Point Cloud Correspondences with TransformersZi Jian Yew, Gim Hee LeeCVPR 2022 · 242 citations
- Generative 3D Part Assembly via Dynamic Graph LearningGuanqi Zhan, Qingnan Fan, Kaichun Mo, Lin Shao et al.NeurIPS 2020 · 113 citations
- Learning Part Generation and Assembly for Structure-Aware Shape SynthesisJun Li, Chengjie Niu, Kai XuAAAI 2020 · 85 citations
- CompoNet: Learning to Generate the Unseen by Part Synthesis and CompositionNadav Schor, Oren Katzir, Hao Zhang, Daniel Cohen-OrICCV 2019 · 63 citations
- Jigsaw: Learning to Assemble Multiple Fractured ObjectsJiaxin Lu, Yifan Sun, Qixing HuangNeurIPS 2023 · 44 citations
Related papers
- Beyond Reassembly: Fractured Object Recovery with Missing PartsQun-Ce Xu, Jiahui Li, Yan-Pei Cao, Weihao Cheng et al.CVPR 2026
- Multiple View Geometry Transformers for 3D Human Pose EstimationZiwei Liao, Jialiang Zhu, Chunyu Wang, Han Hu et al.CVPR 2024
- HiPART: Hierarchical Pose AutoRegressive Transformer for Occluded 3D Human Pose EstimationHongwei Zheng, Han Li, Wenrui Dai, Ziyang Zheng et al.CVPR 2025
- 3D Geometric Shape Assembly via Efficient Point Cloud MatchingNahyuk Lee, Juhong Min, Junha Lee, Seungwook Kim et al.ICML 2024 · 12 citations
- Deep Semantic Graph Transformer for Multi-View 3D Human Pose EstimationLijun Zhang, Kangkang Zhou, Feng Lu, Xiang-Dong Zhou et al.AAAI 2024 · 14 citations
