RenderFormer: Transformer-based Neural Rendering of Triangle Meshes with Global Illumination
Chong Zeng, Yue Dong, Pieter Peers, Hongzhi Wu, Xin Tong
Abstract
We present RenderFormer, a neural rendering pipeline that directly renders an image from a triangle-based representation of a scene with full global illumination effects and that does not require per-scene training or fine-tuning. Instead of taking a physics-centric approach to rendering, we formulate rendering as a sequence-to-sequence transformation where a sequence of tokens representing triangles with reflectance properties is converted to a sequence of output tokens representing small patches of pixels. RenderFormer follows a two stage pipeline: a view-independent stage that models triangle-to-triangle light transport, and a view-dependent stage that transforms a token representing a bundle of rays to the corresponding pixel values guided by the triangle-sequence from the view-independent stage. Both stages are based on the transformer architecture and are learned with minimal prior constraints. We demonstrate and evaluate RenderFormer on scenes with varying complexity in shape and light transport.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6146b168-a4c1-48a9-a3aa-d21510529d35Cited by top-tier papers10
- CUPID: Generative 3D Reconstruction via Joint Object and Pose ModelingBinbin Huang, Haobin Duan, Yiqun Zhao, Zibo Zhao et al.CVPR 2026 · 8 citations
- NeAR: Coupled Neural Asset-Renderer StackHong Li, Chongjie Ye, Houyuan Chen, Weiqing Xiao et al.CVPR 2026 · 4 citations
- RenderFlow: Single-Step Neural Rendering via Flow MatchingShenghao Zhang, Runtao Liu, Christopher Schroers, Yang ZhangCVPR 2026 · 3 citations
- RnG: A Unified Transformer for Complete 3D Modeling from Partial ObservationsMochu Xiang, Zhelun Shen, Xuesong li, Jiahui Ren et al.CVPR 2026 · 2 citations
- HoloPathTracer: Fast and Accurate Wave Path Tracing for HolographyWenbin Zhou, Xiangyu Meng, Jiankai Xing, Xin Liu et al.SIGGRAPH 2026 · 1 citation
Builds on19
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- FlashAttention-2: Faster Attention with Better Parallelism and Work PartitioningTri DaoICLR 2024 · 2,600 citations
- Vision Transformers Need RegistersTimothée Darcet, Maxime Oquab, Julien Mairal, Piotr BojanowskiICLR 2024 · 769 citations
- Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category ReconstructionJeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone et al.ICCV 2021 · 686 citations
Related papers
- A Generalizable Light Transport 3D Embedding for Global IlluminationBing Xu, Mukund Varma T., Cheng Wang, Tzu-Mao Li et al.SIGGRAPH 2026
- LightFormer: Light-Oriented Global Neural Rendering in Dynamic SceneHaocheng Ren, Yuchi Huo, Yifan Peng, Hongtao Sheng et al.SIGGRAPH 2024 · 7 citations
- Is Attention All That NeRF Needs?Mukund Varma T., Peihao Wang, Xuxi Chen, Tianlong Chen et al.ICLR 2023 · 6 citations
- InNeRF: Learning Interpretable Radiance Fields for Generalizable 3D Scene Representation and RenderingDan Wang, Xinrui CuiACM MM 2024
- Scene Representation Transformer: Geometry-Free Novel View Synthesis Through Set-Latent Scene RepresentationsMehdi S. M. Sajjadi, Henning Meyer, Etienne Pot, Urs Bergmann et al.CVPR 2022 · 102 citations
