DeepFormableTag: end-to-end generation and recognition of deformable fiducial markers
Mustafa B. Yaldiz, Andreas Meuleman, Hyeonjoong Jang, Hyunho Ha, Min H. Kim
Abstract
Fiducial markers have been broadly used to identify objects or embed messages that can be detected by a camera. Primarily, existing detection methods assume that markers are printed on ideally planar surfaces. The size of a message or identification code is limited by the spatial resolution of binary patterns in a marker. Markers often fail to be recognized due to various imaging artifacts of optical/perspective distortion and motion blur. To overcome these limitations, we propose a novel deformable fiducial marker system that consists of three main parts: First, a fiducial marker generator creates a set of free-form color patterns to encode significantly large-scale information in unique visual codes. Second, a differentiable image simulator creates a training dataset of photorealistic scene images with the deformed markers, being rendered during optimization in a differentiable manner. The rendered images include realistic shading with specular reflection, optical distortion, defocus and motion blur, color alteration, imaging noise, and shape deformation of markers. Lastly, a trained marker detector seeks the regions of interest and recognizes multiple marker patterns simultaneously via inverse deformation transformation. The deformable marker creator and detector networks are jointly optimized via the differentiable photorealistic renderer in an end-to-end manner, allowing us to robustly recognize a wide range of deformable markers with high accuracy. Our deformable marker system is capable of decoding 36-bit messages successfully at 29 fps with severe shape deformation. Results validate that our system significantly outperforms the traditional and data-driven marker methods. Our learning-based marker system opens up new interesting applications of fiducial markers, including cost-effective motion capture of the human body, active 3D scanning using our fiducial markers' array as structured light patterns, and robust augmented reality rendering of virtual objects on dynamic surfaces.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b64a6967-40c0-4cdd-8610-66172e4c0f67Cited by top-tier papers2
- InfraredTags: Embedding Invisible AR Markers and Barcodes Using Low-Cost, Infrared-Based 3D Printing and Imaging ToolsMustafa Doga Dogan, Ahmad Taka, Michael Lu, Yunyi Zhu et al.CHI 2022 · 59 citations
- Neural Lens ModelingWenqi Xian, Aljaz Bozic, Noah Snavely, Christoph LassnerCVPR 2023
Builds on1
Related papers
- TsFPS: An Accurate and Flexible 6DoF Tracking System with Fiducial Platonic SolidsNan Xiang, Xiaosong Yang, Jian J. ZhangACM MM 2021 · 7 citations
- High-Fidelity 4D Cloth Capture Pipeline with a Two-Level PatternZiheng Liu, Anka He Chen, Shu Chen, Yin Yang et al.SIGGRAPH 2026
- Capturing detailed deformations of moving human bodiesHe Chen, Hyojoon Park, Kutay Macit, Ladislav KavanSIGGRAPH 2021 · 30 citations
- Deep Learning Super-Resolution Network Facilitating Fiducial Tangibles on Capacitive TouchscreensMarius Mihai Rusu, Sven MayerCHI 2023 · 8 citations
- Deep 3D-to-2D Watermarking: Embedding Messages in 3D Meshes and Extracting Them from 2D RenderingsInnfarn Yoo, Huiwen Chang, Xiyang Luo, Ondrej Stava et al.CVPR 2022 · 39 citations
