ORBIT: A Prognostic World Model for Ocular Reasoning Based on Imagined Trajectories
Jiangtao Yan, Yanlin Qu, Yansheng Qiu, Shujian Gao, Wei Yu, Zheng Wang, Xiaodong Sun, Huixun Jia, Diping Song
摘要
The longitudinal management of blinding fundus diseases constitutes a Partially Observable Markov Decision Process (POMDP) necessitating a critical precision-risk trade-off between intervention and over-treatment, as true pathology is often obscured in static observations. However, existing paradigms fail to address this complexity. Traditional vision models remain uninterpretable and memoryless, and while Vision-Language Models (VLMs) excel in semantic understanding, they rely on unsafe open-loop text reasoning lacking the anatomical grounding essential for clinical safety. Furthermore, robust learning is hindered by the scarcity of process supervision in sparse clinical records. To bridge this gap, we introduce the Logic-Constrained Abductive Data Engine. Operating on a ``Propose-and-Verify'' paradigm, it validates MLLM-proposed biomarkers against clinical and temporal logic to reconstruct dense pathological states from sparse outcomes. Building on this foundation, we propose ORBIT, the first ophthalmic Prognostic World Model. Uniquely, ORBIT employs counterfactual visual foresight to imagine anatomical futures under different treatments, anchoring decisions in Closed-Loop Anatomical Verification rather than linguistic probabilities. Experiments demonstrate that ORBIT effectively captures disease evolution and establishes a new paradigm for human-in-the-loop longitudinal decision support and anatomically grounded treatment planning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper16
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Tree of Thoughts: Deliberate Problem Solving with Large Language ModelsShunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran 等NeurIPS 2023 · 被引用 5,068 次
- Self-Refine: Iterative Refinement with Self-FeedbackAman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan 等NeurIPS 2023 · 被引用 4,972 次
- Faith and Fate: Limits of Transformers on CompositionalityNouha Dziri, Ximing Lu, Melanie Sclar, Xiang Lorraine Li 等NeurIPS 2023 · 被引用 728 次
- QoQ-Med: Building Multimodal Clinical Foundation Models with Domain-Aware GRPO TrainingDavid Dai, Peilin Chen, Chanakya Ekbote, Paul Pu LiangNeurIPS 2025 · 被引用 48 次
相关 Paper
- Constructing Ophthalmic MLLM for Positioning-Diagnosis Collaboration Through Clinical Cognitive Chain ReasoningXinyao Liu, Diping SongICCV 2025 · 被引用 9 次
- Medic-AD: Towards Medical Vision-Language Model's Clinical IntelligenceWoohyeon Park, Jaeik Kim, Sunghwan Steve Cho, Pa Hong 等CVPR 2026
- NL-Eye: Abductive NLI For ImagesMor Ventura, Michael Toker, Nitay Calderon, Zorik Gekhman 等ICLR 2025
- EyecareGPT: Boosting Comprehensive Ophthalmology Understanding with Tailored Dataset, Benchmark and ModelSijing Li, Tianwei Lin, Lingshuai Lin, Wenqiao Zhang 等ACM MM 2025 · 被引用 6 次
- Incomplete Modality Disentangled Representation for Ophthalmic Disease Grading and DiagnosisChengzhi Liu, Zile Huang, Zhe Chen, Feilong Tang 等AAAI 2025 · 被引用 10 次
