Copy-Transform-Paste: Zero-Shot Object-Object Alignment Guided by Vision-Language and Geometric Constraints
Rotem Gatenyo, Ohad Fried
2026年份
摘要
Burger bun bottom, lettuce, burger patty, cheese, tomatoes and burger bun top" "Nigiri"
"Captain America holding a shield" "Pinocchio wearing a hat" "A stand with a golden necklace''
Iterative procedure Figure 1. Text-guided object-object alignment and iterative composition. The figure shows four independent examples, each presenting the input meshes and text prompt alongside our alignment result. In addition, an iterative example demonstrates progressive assembly of a burger: the output of stage k is incorporated into the input of stage k+1, gradually forming the final arrangement.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper19
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Scaling Up Visual and Vision-Language Representation Learning With Noisy Text SupervisionChao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen 等ICML 2021 · 被引用 5,401 次
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or 等ICCV 2021 · 被引用 1,437 次
- Soft Rasterizer: A Differentiable Renderer for Image-Based 3D ReasoningShichen Liu, Weikai Chen, Tianye Li, Hao LiICCV 2019 · 被引用 789 次
- StyleGAN-NADA: CLIP-guided domain adaptation of image generatorsRinon Gal, Or Patashnik, Haggai Maron, Amit H. Bermano 等SIGGRAPH 2022 · 被引用 501 次
相关 Paper
- Latent-NeRF for Shape-Guided Generation of 3D Shapes and TexturesGal Metzer, Elad Richardson, Or Patashnik, Raja Giryes 等CVPR 2023
- The Chosen One: Consistent Characters in Text-to-Image Diffusion ModelsOmri Avrahami, Amir Hertz, Yael Vinker, Moab Arar 等SIGGRAPH 2024 · 被引用 26 次
- ArtFormer: Controllable Generation of Diverse 3D Articulated ObjectsJiayi Su, Youhe Feng, Zheng Li, Jinhua Song 等CVPR 2025
- PuzzleFusion++: Auto-agglomerative 3D Fracture Assembly by Denoise and VerifyZhengqing Wang, Jiacheng Chen, Yasutaka FurukawaICLR 2025
- TAPS3D: Text-Guided 3D Textured Shape Generation from Pseudo SupervisionJiacheng Wei, Hao Wang, Jiashi Feng, Guosheng Lin 等CVPR 2023
