Presenting a Paper is an Art: Self-Improvement Aesthetic Agents for Academic Presentations
Chengzhi Liu, Yuzhe YANG, Kaiwen Zhou, Zhen Zhang, Yue Fan, Yanan Xie, Peng Qi, Xin Eric Wang
摘要
The promotion of academic papers has become an important means of enhancing research visibility. where the appeal of dissemination largely determines its effectiveness. However, existing automated methods struggle limited storytelling, insufficient aesthetic quality, and constrained self-adjustment, making it difficult to achieve efficient and engaging dissemination. At the heart of those challenges is a simple principle: there is no way to improve it when you cannot evaluate it right. To address this, we introduce EvoPresent, a self-improvement agent framework that unifies coherent narratives, aesthetic-aware designs, and realistic presentation delivery via virtual characters. Central to EvoPresent is PresAesth, a multi-task reinforcement learning (RL) aesthetic model that provides reliable aesthetic scoring, defect adjustment, and comparative feedback, enabling iterative self-improvement even under limited aesthetic training data. To systematically evaluate the methods, we introduce EvoPresent Benchmark, a comprehensive benchmark comprising: Presentation Generation Quality, built on 650 top-tier AI conference papers with multimodal resources (slides, videos and scripts) to assess both content and design; and Aesthetic Awareness, consisting of 2,000 slide pairs with varying aesthetic levels, supporting joint training and evaluation on scoring, defect adjustment, and comparison. Our findings highlight that (i) High-quality feedback is essential for agent self-improvement, while initial capability alone does not guarantee effective self-correction. (ii) Automated generation pipelines exhibit a trade-off between visual design and content construction. (iii) Multi-task RL training shows stronger generalization in aesthetic awareness tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning ModelsZhongxing Xu, Chengzhi Liu, Qingyue Wei, Juncheng Wu 等NeurIPS 2025 · 被引用 103 次
- Reasoning over Precedents Alongside Statutes: Case-Augmented Deliberative Alignment for LLM SafetyCan Jin, Rui Wu, Tong Che, Qixin Zhang 等ACL 2026 · 被引用 3 次
它引用的顶会 Paper8
- Q-Insight: Understanding Image Quality via Visual Reinforcement LearningWeiqi Li, Xuanyu Zhang, Shijie Zhao, Yabin Zhang 等NeurIPS 2025 · 被引用 117 次
- FLOAT: Generative Motion Latent Flow Matching for Audio-Driven Talking PortraitTaekyung Ki, Dongchan Min, Gyeongsu ChaeICCV 2025 · 被引用 6 次
- Fooling the LVLM Judges: Visual Biases in LVLM-Based EvaluationYerin Hwang, Dongryeol Lee, Kyungmin Min, Taegwan Kang 等EMNLP 2025 · 被引用 4 次
- PPTAgent: Generating and Evaluating Presentations Beyond Text-to-SlidesHao Zheng, Xinyan Guan, Hao Kong, Wenkai Zhang 等EMNLP 2025 · 被引用 3 次
- PhyT2V: LLM-Guided Iterative Self-Refinement for Physics-Grounded Text-to-Video GenerationQiyao Xue, Xiangyu Yin, Boyuan Yang, Wei GaoCVPR 2025
相关 Paper
- P2P: Automated Paper-to-Poster Generation and Fine-Grained BenchmarkTao Sun, Enhao Pan, Zhengkai Yang, Kaixin Sui 等ICLR 2026 · 被引用 19 次
- Code Aesthetics with Agentic Reward FeedbackBang Xiao, Lingjie Jiang, Shaohan Huang, Tengchao Lv 等ICLR 2026 · 被引用 7 次
- PlotThread: Creating Expressive Storyline Visualizations using Reinforcement LearningTan Tang, Renzhong Li, Xinke Wu, Shuhan Liu 等IEEE VIS 2020 · 被引用 70 次
- PPT-Eval: A Benchmark for Computer-Use Agents on PowerPoint TasksApurva Gandhi, Vishwas Suryanarayanan, Raja Anwar, Firoz Shaik 等ICML 2026 · 被引用 3 次
- ReviewRL: Towards Automated Scientific Review with RLSihang Zeng, Kai Tian, Kaiyan Zhang, Yuru Wang 等EMNLP 2025
