Eval3D: Interpretable and Fine-grained Evaluation for 3D Generation
Shivam Duggal, Yushi Hu, Oscar Michel, Aniruddha Kembhavi, William T. Freeman, Noah A. Smith, Ranjay Krishna, Antonio Torralba, Ali Farhadi, Wei-Chiu Ma
Abstract
A plate of fried chicken and waffles with maple syrup A corgi wearing a hat front back side front side Structural inconsistency / Text-3D misalignment / Semantic inconsistency / Geometric inconsistency Figure 1. Challenges of 3D generation: (1) Structural inconsistency: lack of globally-coherent 3D shape; (2) Text-3D misalignment: failure to meet the requirements of the input text-prompt; (3) Semantic inconsistency: content change and incoherent semantics; (4) Geometric inconsistency: misaligned geometry and texture.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c2931758-7637-4a6f-b2b5-77b72cc6ccddCited by top-tier papers2
- 4DWorldBench: A Comprehensive Evaluation Framework for 3D/4D World Generation ModelsYiting Lu, Wei Luo, Peiyan Tu, Haoran Li et al.CVPR 2026 · 10 citations
- GuideFlow3D: Optimization-Guided Rectified Flow For Appearance TransferSayan Deb Sarkar, Sinisa Stekovic, Vincent Lepetit, Iro ArmeniNeurIPS 2025 · 3 citations
Builds on30
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt et al.NeurIPS 2021 · 2,500 citations
Related papers
- DiffSplat: Repurposing Image Diffusion Models for Scalable Gaussian Splat GenerationChenguo Lin, Panwang Pan, Bangbang Yang, Zeming Li et al.ICLR 2025
- Enhancing 3D Fidelity of Text-to-3D using Cross-View CorrespondencesSeungwook Kim, Kejie Li, Xueqing Deng, Yichun Shi et al.CVPR 2024
- SweetDreamer: Aligning Geometric Priors in 2D diffusion for Consistent Text-to-3DWeiyu Li, Rui Chen, Xuelin Chen, Ping TanICLR 2024 · 155 citations
- Diffusion Feature Field for Text-based 3D Editing with Gaussian SplattingEunseo Koh, Sangeek Hyun, MinKyu Lee, Jiwoo Chung et al.NeurIPS 2025 · 5 citations
- Check, Locate, Rectify: A Training-Free Layout Calibration System for Text- to- Image GenerationBiao Gong, Siteng Huang, Yutong Feng, Shiwei Zhang et al.CVPR 2024
