Lune

CVPR2026顶会

SAT-RRG: LLM-Guided Self-Adaptive Training for Radiology Report Generation with Token-Level Push-Pull Optimization

Yunyi Liu, Yingshu Li, Tong Chen, Lingqiao Liu, Lei Wang, Luping Zhou

出版方
2026年份

摘要

Radiology report generators often produce fluent text yet miss crucial details, leading to local semantic conflicts or flipped findings that require stronger penalties. Crossentropy (CE) merely increases the probability of the ground-truth token y * without directly suppressing the model's current wrong choice ŷ, and treats all positions uniformly, so corrections are not prioritized. We introduce a self-adaptive optimization framework that dynamically adjusts token-level gradients based on semantic discrepancy cues derived from a frozen LLM referee. The LLM itself is not the contribution-it merely provides weak supervision to trigger the adaptive learning process. Within this framework, (i) semantic conflicts between the predicted and reference reports are automatically localized and tagged with <e>...</e> (used only during training), and (ii) adaptive, stronger penalties are applied within these sparse but critical spans. Updates follow a push-pull scheme: error spans are pushed down, while non-error tokens are reinforced. The update strength is governed by two complementary signals-normalized entropy (for uncertainty calibration) and focal-style confidence (for handling overand under-confident predictions). On MIMIC-CXR and IU-Xray, our framework consistently improves both language metrics (BLEU-4, ROUGE-L, METEOR) and clinical metrics (RadGraph F1, CheXbert), and remains robust to noisy or imperfect error tags.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

它引用的顶会 Paper18

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖