RealTwin: Concept Graph Representation and Grounding Framework for Reality-Preserving Digital Twin Reconstruction
Zisu Li, Ruohao Li, Jiawei Li, Chao Liu, Junyi Zhu, Daniela Rus, Chen Liang, Mingming Fan
摘要
Reconstructing realistic digital twins has become crucial as advances in mixed reality, metaverse, and robotics demand more accurate simulations for the physical world. Despite technical progress, building high-fidelity digital twins from a systematic and human-centered perspective remains underexplored. Drawing from the human processing model, we decompose human-centric reality into perception, motion, and cognition, and define a reality-preserving digital twin (RPDT) as a reconstruction integrating these dimensions. We present RealTwin, an attribute-graph-based representation and inference framework for RPDT. Leveraging the grounding capabilities of Multimodal Large Language Models (MLLMs), RealTwin chains AI tools to construct attribute graphs that faithfully encode real-world properties. We validate RealTwin through both technical evaluation, showing promising success in graph parsing and attribute inference, and a user study, assessing its applicability across diverse user groups. Enlightened by RealTwin, we discuss critical issues, including ecology, interaction space, and real-world adoption, for future end-to-end, fine-grained, and scalable digital twin reconstruction.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper40
- NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language ModelsGengze Zhou, Yicong Hong, Qi WuAAAI 2024 · 被引用 361 次
- Text2Tex: Text-driven Texture Synthesis via Diffusion ModelsDave Zhenyu Chen, Yawar Siddiqui, Hsin-Ying Lee, Sergey Tulyakov 等ICCV 2023 · 被引用 262 次
- VR-GS: A Physical Dynamics-Aware Interactive Gaussian Splatting System in Virtual RealityYing Jiang, Chang Yu, Tianyi Xie, Xuan Li 等SIGGRAPH 2024 · 被引用 153 次
- Holistic++ Scene Understanding: Single-View 3D Holistic Scene Parsing and Human Pose Estimation With Human-Object Interaction and Physical CommonsenseYixin Chen, Siyuan Huang, Tao Yuan, Yixin Zhu 等ICCV 2019 · 被引用 130 次
- RigNet: neural rigging for articulated charactersZhan Xu, Yang Zhou, Evangelos Kalogerakis, Chris Landreth 等SIGGRAPH 2020 · 被引用 127 次
相关 Paper
- ArtLLM: Generating Articulated Assets via 3D LLMPenghao Wang, Siyuan Xie, Hongyu Yan, Xianghui Yang 等CVPR 2026 · 被引用 7 次
- URDF-Anything: Constructing Articulated Objects with 3D Multimodal Language ModelZhe Li, Xiang Bai, Jieyu Zhang, Zhuangzhe Wu 等NeurIPS 2025 · 被引用 24 次
- Continuously Updating Digital Twins using Large Language ModelsHarry Amad, Nicolás Astorga, Mihaela van der SchaarICML 2025
- Online Reasoning Video Segmentation with Just-in-Time Digital TwinsYiqing Shen, Bohan Liu, Chenjia Li, Lalithkumar Seenivasan 等ICCV 2025 · 被引用 7 次
- Ditto: Building Digital Twins of Articulated Objects from InteractionZhenyu Jiang, Cheng-Chun Hsu, Yuke ZhuCVPR 2022 · 被引用 77 次
