Refracting Reality: Generating Images with Realistic Transparent Objects
Yue Yin, Enze Tao, Dylan Campbell
摘要
Generative image models can produce convincingly real images, with plausible shapes, textures, layouts and lighting. However, one domain in which they perform notably poorly is in the synthesis of transparent objects, which exhibit refraction, reflection, absorption and scattering. Refraction is a particular challenge, because refracted pixel rays often intersect with surfaces observed in other parts of the image, providing a constraint on the color. It is clear from inspection that generative models have not distilled the laws of optics sufficiently well to accurately render refractive objects. In this work, we consider the problem of generating images with accurate refraction, given a text prompt. We synchronize the pixels within the object's boundary with those outside by warping and merging the pixels using Snell's Law of Refraction, at each step of the generation trajectory. For those surfaces that are not directly observed in the image, but are visible via refraction or reflection, we recover their appearance by synchronizing the image with a second generated image---a panorama centered at the object---using the same warping and merging procedure. We demonstrate that our approach generates much more optically-plausible images that respect the physical constraints.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper38
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann 等ICLR 2024 · 被引用 4,569 次
- Scaling Rectified Flow Transformers for High-Resolution Image SynthesisPatrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari 等ICML 2024 · 被引用 3,620 次
- ImageReward: Learning and Evaluating Human Preferences for Text-to-Image GenerationJiazheng Xu, Xiao Liu, Yuchen Wu, Yuxuan Tong 等NeurIPS 2023 · 被引用 1,310 次
相关 Paper
- Eikonal Fields for Refractive Novel-View SynthesisMojtaba Bemana, Karol Myszkowski, Jeppe Revall Frisvad, Hans-Peter Seidel 等SIGGRAPH 2022 · 被引用 38 次
- Differentiable Neural Surface Refinement for Modeling Transparent ObjectsWeijian Deng, Dylan Campbell, Chunyi Sun, Shubham Kanitkar 等CVPR 2024
- NEMTO: Neural Environment Matting for Novel View and Relighting Synthesis of Transparent ObjectsDongqing Wang, Tong Zhang, Sabine SüsstrunkICCV 2023 · 被引用 22 次
- Distortion-free Mid-air Image Inside Refractive Surface and on Reflective SurfaceShunji Kiuchi, Naoya KoizumiIEEE VR 2022 · 被引用 3 次
- Through the Looking Glass: Neural 3D Reconstruction of Transparent ShapesZhengqin Li, Yu-Ying Yeh, Manmohan ChandrakerCVPR 2020
