What Makes A Good Story? Designing Composite Rewards for Visual Storytelling
Junjie Hu, Yu Cheng, Zhe Gan, Jingjing Liu, Jianfeng Gao, Graham Neubig
摘要
Previous storytelling approaches mostly focused on optimizing traditional metrics such as BLEU, ROUGE and CIDEr. In this paper, we re-examine this problem from a different angle, by looking deep into what defines a natural and topically-coherent story. To this end, we propose three assessment criteria: relevance, coherence and expressiveness, which we observe through empirical analysis could constitute a “high-quality” story to the human eye. We further propose a reinforcement learning framework, ReCo-RL, with reward functions designed to capture the essence of these quality criteria. Experiments on the Visual Storytelling Dataset (VIST) with both automatic and human evaluation demonstrate that our ReCo-RL model achieves better performance than state-of-the-art baselines on both traditional metrics and the proposed new criteria.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Commonsense Knowledge Aware Concept Selection For Diverse and Informative Visual StorytellingHong Chen, Yifei Huang, Hiroya Takamura, Hideki NakayamaAAAI 2021 · 被引用 49 次
- Imagine, Reason and Write: Visual Storytelling with Graph Knowledge and Relational ReasoningChunpu Xu, Min Yang, Chengming Li, Ying Shen 等AAAI 2021 · 被引用 39 次
- AESOP: Abstract Encoding of Stories, Objects, and PicturesHareesh Ravi, Kushal Kafle, Scott Cohen, Jonathan Brandt 等ICCV 2021 · 被引用 19 次
- Latent Memory-augmented Graph Transformer for Visual StorytellingMengshi Qi, Jie Qin, Di Huang, Zhiqiang Shen 等ACM MM 2021 · 被引用 18 次
- Topic Adaptation and Prototype Encoding for Few-Shot Visual StorytellingJiacheng Li, Siliang Tang, Juncheng Li, Jun Xiao 等ACM MM 2020 · 被引用 9 次
相关 Paper
- Learning to Rank Visual Stories From Human Ranking DataChi-Yang Hsu, Yun-Wei Chu, Vincent Chen, Kuan-Chieh Lo 等ACL 2022
- Improving Multi-party Dialogue Generation via Topic and Rhetorical CoherenceYaxin Fan, Peifeng Li, Qiaoming ZhuEMNLP 2024 · 被引用 1 次
- Triangle-Reward Reinforcement Learning: A Visual-Linguistic Semantic Alignment for Image CaptioningWeizhi Nie, Jiesi Li, Ning Xu, An-An Liu 等ACM MM 2021 · 被引用 9 次
- Evaluating Visual Narrative Coherence in Story Visualization via Diversified StorylinesMinha Jhang, Kyeongman Park, Hyukhun Koh, Kyomin JungACL 2026
- PaCo-RL: Advancing Reinforcement Learning for Consistent Image Generation with Pairwise Reward ModelingBowen Ping, Chengyou Jia, Minnan Luo, Changliang Xia 等CVPR 2026 · 被引用 9 次
