CLER: Improving Multimodal Financial Reasoning by Cross-MLLM Error Reflection
Shuangyan Deng, Zhongsheng Wang, Rui Mao, Ciprian Doru Giurcaneanu, Jiamou Liu
摘要
Recent advances in Multimodal Large Language Models (MLLMs) have enabled joint reasoning over financial textual and visual inputs. However, they still struggle with financial terminology, logical consistency, and numerical computations. Moreover, while commercial large models perform well on reasoning tasks, their high inference costs limit their scalable usage in real world financial applications. We thus propose a cost-effective framework, CLER, that combines contrastive retrieval with step-wise reflection to improve reasoning performance. Also, the reasoning cost is only generated in the test stage when using commercial large models. CLER leverages FinErrorSet, a dataset of 8,000+ mistake correction pairs from diverse open-source MLLMs. A fine grained retriever is trained to identify structurally relevant errors for self-correction through individual reflection. Experiments on three benchmarks show that CLER consistently outperforms other baselines. To our knowledge, CLER is the first framework to use cross-model errors for financial reasoning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- MMMU: A Massive Multi-Discipline Multimodal Understanding and Reasoning Benchmark for Expert AGIXiang Yue, Yuansheng Ni, Tianyu Zheng, Kai Zhang 等CVPR 2024 · 被引用 213 次
- Can Language Models Perform Robust Reasoning in Chain-of-thought Prompting with Noisy Rationales?Zhanke Zhou, Rong Tao, Jianing Zhu, Yiwen Luo 等NeurIPS 2024 · 被引用 74 次
- Multimodal Multi-Task Financial Risk ForecastingRamit Sawhney, Puneet Mathur, Ayush Mangal, Piyush Khanna 等ACM MM 2020 · 被引用 61 次
- In-Context Principle Learning from MistakesTianjun Zhang, Aman Madaan, Luyu Gao, Steven Zheng 等ICML 2024 · 被引用 44 次
相关 Paper
- FinMMDocR: Benchmarking Financial Multimodal Reasoning with Scenario Awareness, Document Understanding, and Multi-Step ComputationZichen Tang, Haihong E, Rongjin Li, Jiacheng Liu 等AAAI 2026
- FCMR: Robust Evaluation of Financial Cross-Modal Multi-Hop ReasoningSeunghee Kim, Changhyeon Kim, Taeuk KimACL 2025
- FinMMR: Make Financial Numerical Reasoning More Multimodal, Comprehensive, and ChallengingZichen Tang, Haihong E, Jiacheng Liu, Zhongjun Yang 等ICCV 2025 · 被引用 1 次
- Retrieval Enhanced Feedback via In-context Neural Error-bookJongyeop Hyun, Bumsoo KimEMNLP 2025
- Program of Thoughts for Financial Reasoning: Leveraging Dynamic In-Context Examples and Generative RetrievalSubhendu Khatuya, Shashwat Naidu, Pawan Goyal, Niloy GangulyEMNLP 2025 · 被引用 3 次
