ORBIT: A Prognostic World Model for Ocular Reasoning Based on Imagined Trajectories
Jiangtao Yan, Yanlin Qu, Yansheng Qiu, Shujian Gao, Wei Yu, Zheng Wang, Xiaodong Sun, Huixun Jia, Diping Song
Abstract
The longitudinal management of blinding fundus diseases constitutes a Partially Observable Markov Decision Process (POMDP) necessitating a critical precision-risk trade-off between intervention and over-treatment, as true pathology is often obscured in static observations. However, existing paradigms fail to address this complexity. Traditional vision models remain uninterpretable and memoryless, and while Vision-Language Models (VLMs) excel in semantic understanding, they rely on unsafe open-loop text reasoning lacking the anatomical grounding essential for clinical safety. Furthermore, robust learning is hindered by the scarcity of process supervision in sparse clinical records. To bridge this gap, we introduce the Logic-Constrained Abductive Data Engine. Operating on a ``Propose-and-Verify'' paradigm, it validates MLLM-proposed biomarkers against clinical and temporal logic to reconstruct dense pathological states from sparse outcomes. Building on this foundation, we propose ORBIT, the first ophthalmic Prognostic World Model. Uniquely, ORBIT employs counterfactual visual foresight to imagine anatomical futures under different treatments, anchoring decisions in Closed-Loop Anatomical Verification rather than linguistic probabilities. Experiments demonstrate that ORBIT effectively captures disease evolution and establishes a new paradigm for human-in-the-loop longitudinal decision support and anatomically grounded treatment planning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 18e1c9dc-f83d-4672-af30-57fe80b9caefBuilds on16
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Tree of Thoughts: Deliberate Problem Solving with Large Language ModelsShunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran et al.NeurIPS 2023 · 5,068 citations
- Self-Refine: Iterative Refinement with Self-FeedbackAman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan et al.NeurIPS 2023 · 4,972 citations
- Faith and Fate: Limits of Transformers on CompositionalityNouha Dziri, Ximing Lu, Melanie Sclar, Xiang Lorraine Li et al.NeurIPS 2023 · 728 citations
- QoQ-Med: Building Multimodal Clinical Foundation Models with Domain-Aware GRPO TrainingDavid Dai, Peilin Chen, Chanakya Ekbote, Paul Pu LiangNeurIPS 2025 · 48 citations
Related papers
- Constructing Ophthalmic MLLM for Positioning-Diagnosis Collaboration Through Clinical Cognitive Chain ReasoningXinyao Liu, Diping SongICCV 2025 · 9 citations
- Medic-AD: Towards Medical Vision-Language Model's Clinical IntelligenceWoohyeon Park, Jaeik Kim, Sunghwan Steve Cho, Pa Hong et al.CVPR 2026
- NL-Eye: Abductive NLI For ImagesMor Ventura, Michael Toker, Nitay Calderon, Zorik Gekhman et al.ICLR 2025
- EyecareGPT: Boosting Comprehensive Ophthalmology Understanding with Tailored Dataset, Benchmark and ModelSijing Li, Tianwei Lin, Lingshuai Lin, Wenqiao Zhang et al.ACM MM 2025 · 6 citations
- Incomplete Modality Disentangled Representation for Ophthalmic Disease Grading and DiagnosisChengzhi Liu, Zile Huang, Zhe Chen, Feilong Tang et al.AAAI 2025 · 10 citations
