Multi-Object Manipulation via Object-Centric Neural Scattering Functions
Stephen Tian, Yancheng Cai, Hong-Xing Yu, Sergey Zakharov, Katherine Liu, Adrien Gaidon, Yunzhu Li, Jiajun Wu
摘要
Learned visual dynamics models have proven effective for robotic manipulation tasks. Yet, it remains unclear how best to represent scenes involving multi-object interactions. Current methods decompose a scene into discrete objects, but they struggle with precise modeling and manipulation amid challenging lighting conditions as they only encode appearance tied with specific illuminations. In this work, we propose using object-centric neural scattering functions (OSFs) as object representations in a model-predictive control framework. OSFs model per-object light transport, enabling compositional scene re-rendering under object rearrangement and varying lighting conditions. By combining this approach with inverse parameter estimation and graph-based neural dynamics models, we demonstrate improved model-predictive control performance and generalization in compositional multi-object environments, even in previously unseen scenarios and harsh lighting conditions. * indicates equal contribution. Yancheng is affiliated with Fudan University; this work was done while he was a summer intern at Stanford.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Inferring Hybrid Neural Fluid Fields from VideosHong-Xing Yu, Yang Zheng, Yuan Gao, Yitong Deng 等NeurIPS 2023 · 被引用 36 次
- DEL: Discrete Element Learner for Learning 3D Particle Dynamics with Neural RenderingJiaxu Wang, Jingkai Sun, Ziyi Zhang, Junhao He 等NeurIPS 2024 · 被引用 5 次
- Digital Twin Catalog: A Large-Scale Photorealistic 3D Object Digital Twin DatasetZhao Dong, Ka Chen, Zhaoyang Lv, Hong-Xing Yu 等CVPR 2025
- Do Computer Vision Foundation Models Learn the Low-level Characteristics of the Human Visual System?Yancheng Cai, Fei Yin, Dounia Hammou, Rafal MantiukCVPR 2025
它引用的顶会 Paper14
- Dream to Control: Learning Behaviors by Latent ImaginationDanijar Hafner, Timothy P. Lillicrap, Jimmy Ba, Mohammad NorouziICLR 2020 · 被引用 1,852 次
- Mastering Atari with Discrete World ModelsDanijar Hafner, Timothy P. Lillicrap, Mohammad Norouzi, Jimmy BaICLR 2021 · 被引用 1,170 次
- KiloNeRF: Speeding up Neural Radiance Fields with Thousands of Tiny MLPsChristian Reiser, Songyou Peng, Yiyi Liao, Andreas GeigerICCV 2021 · 被引用 963 次
- Pix2Pose: Pixel-Wise Coordinate Regression of Objects for 6D Pose EstimationKiru Park, Timothy Patten, Markus VinczeICCV 2019 · 被引用 527 次
- DPOD: 6D Pose Object Detector and RefinerSergey Zakharov, Ivan Shugurov, Slobodan IlicICCV 2019 · 被引用 486 次
相关 Paper
- Object-Centric Representation Learning with Generative Spatial-Temporal FactorizationNanbo Li, Muhammad Ahmed Raza, Wenbin Hu, Zhaole Sun 等NeurIPS 2021 · 被引用 17 次
- DynaVol: Unsupervised Learning for Dynamic Scenes through Object-Centric VoxelizationYanpeng Zhao, Siyu Gao, Yunbo Wang, Xiaokang YangICLR 2024 · 被引用 2 次
- LightFormer: Light-Oriented Global Neural Rendering in Dynamic SceneHaocheng Ren, Yuchi Huo, Yifan Peng, Hongtao Sheng 等SIGGRAPH 2024 · 被引用 7 次
- Neural Scene Graphs for Dynamic ScenesJulian Ost, Fahim Mannan, Nils Thuerey, Julian Knodt 等CVPR 2021
- DM-NeRF: 3D Scene Geometry Decomposition and Manipulation from 2D ImagesBing Wang, Lu Chen, Bo YangICLR 2023 · 被引用 26 次
