PhysX-3D: Physical-Grounded 3D Asset Generation
Ziang Cao, Zhaoxi Chen, Liang Pan, Ziwei Liu
摘要
3D modeling is moving from virtual to physical. Existing 3D generation primarily emphasizes geometries and textures while neglecting physical-grounded modeling. Consequently, despite the rapid development of 3D generative models, the synthesized 3D assets often overlook rich and important physical properties, hampering their real-world application in physical domains like simulation and embodied AI. As an initial attempt to address this challenge, we propose PhysX-3D, an end-to-end paradigm for physical-grounded 3D asset generation. 1) To bridge the critical gap in physics-annotated 3D datasets, we present PhysXNet - the first physics-grounded 3D dataset systematically annotated across five foundational dimensions: absolute scale, material, affordance, kinematics, and function description. In particular, we devise a scalable human-in-the-loop annotation pipeline based on vision-language models, which enables efficient creation of physics-first assets from raw 3D assets.2) Furthermore, we propose PhysXGen, a feed-forward framework for physics-grounded image-to-3D asset generation, injecting physical knowledge into the pre-trained 3D structural space. Specifically, PhysXGen employs a dual-branch architecture to explicitly model the latent correlations between 3D structures and physical properties, thereby producing 3D assets with plausible physical predictions while preserving the native geometry quality. Extensive experiments validate the superior performance and promising generalization capability of our framework. All the code, data, and models will be released to facilitate future research in generative physical AI.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- World-In-World: World Models in a Closed-Loop WorldJiahan Zhang, Muqing Jiang, Nanru Dai, Taiming Lu 等ICLR 2026 · 被引用 46 次
- PhysX-Anything: Simulation-Ready Physical 3D Assets from Single ImageZiang Cao, Fangzhou Hong, Zhaoxi Chen, Liang Pan 等CVPR 2026 · 被引用 36 次
- ReconViaGen: Towards Accurate Multi-view 3D Object Reconstruction via GenerationJiahao Chang, Chongjie Ye, Yushuang Wu, Yuantao Chen 等ICLR 2026 · 被引用 30 次
- VoMP: Predicting Volumetric Mechanical Property FieldsRishit Dagli, Donglai Xiang, Vismay Modi, Charles Loop 等ICLR 2026 · 被引用 13 次
- PhysGM: Large Physical Gaussian Model for Feed-Forward 4D SynthesisChunji Lv, Zequn Chen, Donglin Di, Weinan Zhang 等CVPR 2026 · 被引用 9 次
它引用的顶会 Paper17
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- DreamFusion: Text-to-3D using 2D DiffusionBen Poole, Ajay Jain, Jonathan T. Barron, Ben MildenhallICLR 2023 · 被引用 463 次
- PhysGaussian: Physics-Integrated 3D Gaussians for Generative DynamicsTianyi Xie, Zeshun Zong, Yuxing Qiu, Xuan Li 等CVPR 2024 · 被引用 118 次
- ABO: Dataset and Benchmarks for Real-World 3D Object UnderstandingJasmine Collins, Shubham Goel, Kenan Deng, Achleshwar Luthra 等CVPR 2022 · 被引用 117 次
- Large-Vocabulary 3D Diffusion Model with TransformerZiang Cao, Fangzhou Hong, Tong Wu, Liang Pan 等ICLR 2024 · 被引用 54 次
相关 Paper
- PhysForge: Generating Physics-Grounded 3D Assets for Interactive Virtual WorldYunhan Yang, Chunshi Wang, Junliang Ye, YANG LI 等ICML 2026 · 被引用 6 次
- PhysInOne: Visual Physics Learning and Reasoning in One SuiteSiyuan Zhou, Hejun Wang, Hu Cheng, Jinxi Li 等CVPR 2026 · 被引用 9 次
- PAT3D: Physics-Augmented Text-to-3D Scene GenerationGuying Lin, Kemeng Huang, Michael Liu, Ruihan Gao 等ICLR 2026 · 被引用 14 次
- AffordGrasp: Cross-Modal Diffusion for Affordance-Aware Grasp SynthesisXiaofei Wu, Yi Zhang, Yumeng Liu, Yuexin Ma 等CVPR 2026 · 被引用 1 次
- PhyCo: Learning Controllable Physical Priors for Generative MotionSriram Narayanan, Ziyu Jiang, Srinivasa G. Narasimhan, Manmohan ChandrakerCVPR 2026 · 被引用 6 次
