DeepFormableTag: end-to-end generation and recognition of deformable fiducial markers
Mustafa B. Yaldiz, Andreas Meuleman, Hyeonjoong Jang, Hyunho Ha, Min H. Kim
摘要
Fiducial markers have been broadly used to identify objects or embed messages that can be detected by a camera. Primarily, existing detection methods assume that markers are printed on ideally planar surfaces. The size of a message or identification code is limited by the spatial resolution of binary patterns in a marker. Markers often fail to be recognized due to various imaging artifacts of optical/perspective distortion and motion blur. To overcome these limitations, we propose a novel deformable fiducial marker system that consists of three main parts: First, a fiducial marker generator creates a set of free-form color patterns to encode significantly large-scale information in unique visual codes. Second, a differentiable image simulator creates a training dataset of photorealistic scene images with the deformed markers, being rendered during optimization in a differentiable manner. The rendered images include realistic shading with specular reflection, optical distortion, defocus and motion blur, color alteration, imaging noise, and shape deformation of markers. Lastly, a trained marker detector seeks the regions of interest and recognizes multiple marker patterns simultaneously via inverse deformation transformation. The deformable marker creator and detector networks are jointly optimized via the differentiable photorealistic renderer in an end-to-end manner, allowing us to robustly recognize a wide range of deformable markers with high accuracy. Our deformable marker system is capable of decoding 36-bit messages successfully at 29 fps with severe shape deformation. Results validate that our system significantly outperforms the traditional and data-driven marker methods. Our learning-based marker system opens up new interesting applications of fiducial markers, including cost-effective motion capture of the human body, active 3D scanning using our fiducial markers' array as structured light patterns, and robust augmented reality rendering of virtual objects on dynamic surfaces.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- InfraredTags: Embedding Invisible AR Markers and Barcodes Using Low-Cost, Infrared-Based 3D Printing and Imaging ToolsMustafa Doga Dogan, Ahmad Taka, Michael Lu, Yunyi Zhu 等CHI 2022 · 被引用 59 次
- Neural Lens ModelingWenqi Xian, Aljaz Bozic, Noah Snavely, Christoph LassnerCVPR 2023
它引用的顶会 Paper1
相关 Paper
- TsFPS: An Accurate and Flexible 6DoF Tracking System with Fiducial Platonic SolidsNan Xiang, Xiaosong Yang, Jian J. ZhangACM MM 2021 · 被引用 7 次
- High-Fidelity 4D Cloth Capture Pipeline with a Two-Level PatternZiheng Liu, Anka He Chen, Shu Chen, Yin Yang 等SIGGRAPH 2026
- Capturing detailed deformations of moving human bodiesHe Chen, Hyojoon Park, Kutay Macit, Ladislav KavanSIGGRAPH 2021 · 被引用 30 次
- Deep Learning Super-Resolution Network Facilitating Fiducial Tangibles on Capacitive TouchscreensMarius Mihai Rusu, Sven MayerCHI 2023 · 被引用 8 次
- Deep 3D-to-2D Watermarking: Embedding Messages in 3D Meshes and Extracting Them from 2D RenderingsInnfarn Yoo, Huiwen Chang, Xiyang Luo, Ondrej Stava 等CVPR 2022 · 被引用 39 次
