Mesh-RFT: Enhancing Mesh Generation via Fine-grained Reinforcement Fine-Tuning
Jian Liu, Jing Xu, Song Guo, Jing Li, Jingfeng Guo, Jiaao Yu, Haohan Weng, Biwen Lei, Xianghui Yang, Zhuo Chen, Fangqi Zhu, Tao Han, Chunchao Guo
摘要
Existing pretrained models for 3D mesh generation often suffer from data biases and produce low-quality results, while global reinforcement learning (RL) methods rely on object-level rewards that struggle to capture local structure details. To address these challenges, we present Mesh-RFT, a novel fine-grained reinforcement fine-tuning framework that employs Masked Direct Preference Optimization (M-DPO) to enable localized refinement via quality-aware face masking. To facilitate efficient quality evaluation, we introduce an objective topology-aware scoring system to evaluate geometric integrity and topological regularity at both object and face levels through two metrics: Boundary Edge Ratio (BER) and Topology Score (TS). By integrating these metrics into a fine-grained RL strategy, Mesh-RFT becomes the first method to optimize mesh quality at the granularity of individual faces, resolving localized errors while preserving global coherence. Experiment results show that our M-DPO approach reduces Hausdorff Distance (HD) by 24.6% and improves Topology Score (TS) by 3.8% over pre-trained models, while outperforming global DPO methods with a 17.4% HD reduction and 4.9% TS gain. These results demonstrate Mesh-RFT's ability to improve geometric integrity and topological regularity, achieving new state-of-the-art performance in production-ready mesh generation. Project Page: https://hitcslj.github.io/mesh-rft/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and UnderstandingJunliang Ye, Zhengyi Wang, Ruowen Zhao, Shenghao Xie 等NeurIPS 2025 · 被引用 42 次
- QuadGPT: Native Quadrilateral Mesh Generation with Autoregressive ModelsJian Liu, Chunshi Wang, Song Guo, Haohan Weng 等ICLR 2026 · 被引用 15 次
- MeshRipple: Structured Autoregressive Generation of Artist-MeshesJunKai Lin, Hang Long, Huipeng Guo, Jielei Zhang 等CVPR 2026 · 被引用 9 次
- LATO: 3D Mesh Flow Matching with Structured TOpology Preserving LAtentsTianhao Zhao, Youjia Zhang, Hang Long, Jinshen Zhang 等ICML 2026 · 被引用 6 次
- FACE: A Face-based Autoregressive Representation for High-Fidelity and Efficient Mesh GenerationHanxiao Wang, Yuanchen Guo, Ying-Tian Liu, Zi-Xin Zou 等CVPR 2026 · 被引用 6 次
它引用的顶会 Paper30
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- DAPO: An Open-Source LLM Reinforcement Learning System at ScaleQiying Yu, Zheng Zhang, Ruofei Zhu, Yufeng Yuan 等NeurIPS 2025 · 被引用 2,828 次
- ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score DistillationZhengyi Wang, Cheng Lu, Yikai Wang, Fan Bao 等NeurIPS 2023 · 被引用 1,498 次
- ImageReward: Learning and Evaluating Human Preferences for Text-to-Image GenerationJiazheng Xu, Xiao Liu, Yuchen Wu, Yuxuan Tong 等NeurIPS 2023 · 被引用 1,310 次
- Efficient Memory Management for Large Language Model Serving with PagedAttentionWoosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng 等SOSP 2023 · 被引用 1,016 次
相关 Paper
- DeepMesh: Auto-Regressive Artist-Mesh Creation with Reinforcement LearningRuowen Zhao, Junliang Ye, Zhengyi Wang, Guangce Liu 等ICCV 2025 · 被引用 9 次
- Mesh-Pro: Asynchronous Advantage-guided Ranking Preference Optimization for Artist-style Quadrilateral Mesh GenerationZhen Zhou, Jian Liu, Biwen Lei, Jing Xu 等CVPR 2026 · 被引用 2 次
- Are We Ready for RL in Text-to-3D Generation? A Progressive InvestigationYiwen Tang, Ziyu Guo, Kaixin Zhu, Ray Zhang 等CVPR 2026 · 被引用 11 次
- BuildingGPT: Auto-Regressive Building Wireframe Reconstruction Model with Reinforcement LearningYuzhou Liu, Lingjie Zhu, Hanqiao Ye, Yujun Liu 等CVPR 2026
- Topology-Preserved Auto-regressive Mesh Generation in the Manner of Weaving SilkGaochao Song, Zibo Zhao, Haohan Weng, Jingbo Zeng 等ICLR 2026 · 被引用 6 次
