DRAWER: Digital Reconstruction and Articulation With Environment Realism
Hongchi Xia, Entong Su, Marius Memmel, Arhan Jain, Raymond Yu, Numfor Mbiziwo-Tiapo, Ali Farhadi, Abhishek Gupta, Shenlong Wang, Wei-Chiu Ma
Abstract
Figure 1. DRAWER automatically converts a video of a static scene into an interactive environment with segmented objects and articulated doors. It supports physical interactions like opening/closing drawers and moving/placing objects, with precise geometry and high-fidelity rendering. With these capabilities, DRAWER can transform a video into interactive games and enable real-to-sim-to-real transfer in robotics.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bb1a6034-89b3-405d-8c9e-97b21d0d9ae0Cited by top-tier papers12
- PointWorld: Scaling 3D World Models for In-The-Wild Robotic ManipulationWenlong Huang, Yu-Wei Chao, Arsalan Mousavian, Ming-Yu Liu et al.CVPR 2026 · 87 citations
- SAGE: Scalable Agentic 3D Scene Generation for Embodied AIHongchi Xia, Xuan Li, Zhaoshuo Li, Qianli Ma et al.CVPR 2026 · 50 citations
- HoloScene: Simulation-Ready Interactive 3D Worlds from a Single VideoHongchi Xia, Chih-Hao Lin, Hao-Yu Hsu, Quentin Leboutet et al.NeurIPS 2025 · 18 citations
- Dexterous World ModelsByungjun Kim, Taeksoo Kim, Junyoung Lee, Hanbyul JooCVPR 2026 · 17 citations
- VoMP: Predicting Volumetric Mechanical Property FieldsRishit Dagli, Donglai Xiang, Vismay Modi, Charles Loop et al.ICLR 2026 · 13 citations
Builds on49
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman et al.ICCV 2021 · 2,700 citations
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt et al.NeurIPS 2021 · 2,500 citations
Related papers
- Video2Game: Real-time, Interactive, Realistic and Browser-Compatible Environment from a Single VideoHongchi Xia, Zhi-Hao Lin, Wei-Chiu Ma, Shenlong WangCVPR 2024 · 9 citations
- DecoupledGaussian: Object-Scene Decoupling for Physics-Based InteractionMiaowei Wang, Yibo Zhang, Weiwei Xu, Rui Ma et al.CVPR 2025
- LeviTor: 3D Trajectory Oriented Image-to-Video SynthesisHanlin Wang, Hao Ouyang, Qiuyu Wang, Wen Wang et al.CVPR 2025
- SketchVideo: Sketch-based Video Generation and EditingFeng-Lin Liu, Hongbo Fu, Xintao Wang, Weicai Ye et al.CVPR 2025
- ChainHOI: Joint-based Kinematic Chain Modeling for Human-Object Interaction GenerationLing-An Zeng, Guohong Huang, Yi-Lin Wei, Shengbo Gu et al.CVPR 2025
