Video-Annotated Augmented Reality Assembly Tutorials
Masahiro Yamaguchi, Shohei Mori, Peter Mohr, Markus Tatzgern, Ana Stanescu, Hideo Saito, Denis Kalkofen
摘要
We present a system for generating and visualizing interactive 3D Augmented Reality tutorials based on 2D video input, which allows viewpoint control at runtime. Inspired by assembly planning, we analyze the input video using a 3D CAD model of the object to determine an assembly graph that encodes blocking relationships between parts. Using an assembly graph enables us to detect assembly steps that are otherwise difficult to extract from the video, and generally improves object detection and tracking by providing prior knowledge about movable parts. To avoid information loss, we combine the 3D animation with relevant parts of the 2D video so that we can show detailed manipulations and tool usage that cannot be easily extracted from the video. To further support user orientation, we visually align the 3D animation with the real-world object by using texture information from the input video. We developed a presentation system that uses commonly available hardware to make our results accessible for home use and demonstrate the effectiveness of our approach by comparing it to traditional video tutorials.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Design Patterns for Situated Visualization in Augmented RealityBenjamin Lee, Michael Sedlmair, Dieter SchmalstiegIEEE VIS 2023 · 被引用 72 次
- InstruMentAR: Auto-Generation of Augmented Reality Tutorials for Operating Digital Instruments Through Recording Embodied DemonstrationZiyi Liu, Zhengzhe Zhu, Enze Jiang, Feichi Huang 等CHI 2023 · 被引用 29 次
- PrISM-Tracker: A Framework for Multimodal Procedure Tracking Using Wearable Sensors and State Transition Information with User-Driven Handling of Errors and UncertaintyRiku Arakawa, Hiromu Yakura, Vimal Mollyn, Suzanne Nie 等UbiComp 2023 · 被引用 20 次
- PrISM-Observer: Intervention Agent to Help Users Perform Everyday Procedures Sensed using a SmartwatchRiku Arakawa, Hiromu Yakura, Mayank GoelUIST 2024 · 被引用 20 次
- VoLearn: A Cross-Modal Operable Motion-Learning System Combined with Virtual Avatar and Auditory FeedbackChengshuo Xia, Xinrui Fang, Riku Arakawa, Yuta SugiuraUbiComp 2022 · 被引用 19 次
相关 Paper
- Task Breakpoint Generation using Origin-Centric Graph in Virtual Reality Recordings for Adaptive PlaybackSelin Choi, Dooyoung Kim, Taewook Ha, Seonji Kim 等IEEE VR 2026
- ARify: Leveraging Narrated Instructional Videos to Create Augmented Reality Tutorials for Procedural TasksXiyun Hu, Chenfei Zhu, Shao-Kang Hsia, Dizhi Ma 等CHI 2026 · 被引用 1 次
- Which Side is the Top? A User Study to Compare Visual Assets for Component Orientation in Assembly with Augmented RealityEnricoandrea Laviola, Michele Gattullo, Sara Romano, Antonio Emmanuele UvaIEEE VR 2025 · 被引用 2 次
- SimRecon: SimReady Compositional Scene Reconstruction from Real VideosChong Xia, Kai Zhu, Zizhuo Wang, Fangfu Liu 等CVPR 2026 · 被引用 11 次
- Neural Assembler: Learning to Generate Fine-Grained Robotic Assembly Instructions from Multi-View ImagesHongyu Yan, Yadong MuAAAI 2025 · 被引用 3 次
