CT-ScanGaze: A Dataset and Baselines for 3D Volumetric Scanpath Modeling
Trong-Thang Pham, Akash Awasthi, Saba Khan, Esteban Duran Marti, Tien-Phat Nguyen, Khoa Vo, Minh Tran, Son Nguyen, Cuong Tran, Yuki Ikebe, Anh Totti Nguyen, Anh Nguyen
摘要
Understanding radiologists' eye movement during Computed Tomography (CT) reading is crucial for developing effective interpretable computer-aided diagnosis systems. However, CT research in this area has been limited by the lack of publicly available eye-tracking datasets and the three-dimensional complexity of CT volumes. To address these challenges, we present the first publicly available eye gaze dataset on CT, called CT-ScanGaze, captured from expert radiologists. Then, we introduce CT-Searcher, a novel 3D scanpath predictor designed specifically to process CT volumes and generate radiologist-like 3D fixation sequences, overcoming the limitations of current scanpath predictors that only handle 2D inputs. Since deep learning models benefit from a pretraining step, we develop a pipeline that converts existing 2D gaze datasets into 3D gaze data to pretrain CT-Searcher. Through both qualitative and quantitative evaluations on CT-ScanGaze, we demonstrate the effectiveness of our approach and provide a comprehensive assessment framework for 3D scanpath prediction in medical imaging. Code and data are available at https://github.com/UARK-AICV/CTScanGaze.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper12
- ViTAE: Vision Transformer Advanced by Exploring Intrinsic Inductive BiasYufei Xu, Qiming Zhang, Jing Zhang, Dacheng TaoNeurIPS 2021 · 被引用 429 次
- UEyes: Understanding Visual Saliency across User Interface TypesYue Jiang, Luis A. Leiva, Hamed Rezazadegan Tavakoli, Paul R. B. Houssel 等CHI 2023 · 被引用 100 次
- Predicting Visual Importance Across Graphic Design TypesCamilo Fosco, Vincent Casser, Amish Kumar Bedi, Peter O'Donovan 等UIST 2020 · 被引用 55 次
- Eye-gaze Guided Multi-modal Alignment for Medical Representation LearningChong Ma, Hanqi Jiang, Wenting Chen, Yiwei Li 等NeurIPS 2024 · 被引用 29 次
- ScanDMM: A Deep Markov Model of Scanpath Prediction for 360° ImagesXiangjie Sui, Yuming Fang, Hanwei Zhu, Shiqi Wang 等CVPR 2023
相关 Paper
- Interpreting Radiologist's Intention from Eye Movements in Chest X-ray DiagnosisTrong-Thang Pham, Anh Nguyen, Zhigang Deng, Carol C. Wu 等ACM MM 2025 · 被引用 1 次
- From Human Attention to Diagnosis: Semantic Patch-Level Integration of Vision-Language Models in Medical ImagingDmitry Lvov, Ilya PershinNeurIPS 2025 · 被引用 2 次
- Mining Gaze for Contrastive Learning toward Computer-Assisted DiagnosisZihao Zhao, Sheng Wang, Qian Wang, Dinggang ShenAAAI 2024 · 被引用 15 次
- VR-DiagNet: Medical Volumetric and Radiomic Diagnosis Networks with Interpretable Clinician-like Optimizing Visual InspectionShouyu Chen, Liang Hu, Tangwei Ye, Zhongyuan Lai 等ACM MM 2024 · 被引用 1 次
- COVID-view: Diagnosis of COVID-19 using Chest CTShreeraj Jadhav, Gaofeng Deng, Marlene Zawin, Arie E. KaufmanIEEE VIS 2021 · 被引用 33 次
