Conditional Visual Autoregressive Modeling for Pathological Image Restoration
Ziyi Liu, Zhe Xu, Jiabo Ma, Wenqiang Li, Ruixuan Wang, Bo Du, Hao Chen
摘要
Pathological image has been recognized as the gold standard for cancer diagnosis for more than a century. However, some internal regions of pathological images may inevitably exhibit various degradation issues, including low resolution, image blurring, and image noising, which will affect disease diagnosis, staging, and risk stratification. Existing pathological image restoration methods were mainly based on generative adversarial networks (GANs) to improve image quality, which are limited by the inherent instability and loss of structural details, often resulting in artifacts in the restored images. Large scale of whole slide images (WSIs) also makes it hard for efficient processing and restoration. To address these limitations, we propose a conditional visual autoregressive model (CVARPath) for next-scale token prediction, guided by the degraded tokens from the current scale. We introduce a novel framework that employs quantified encoders specifically designed for pathological image generation, which learns consistent sparse vocabulary tokens through self-supervised contrastive learning. Furthermore, our method efficiently compresses image patches into compact degraded sparse tokens at smaller scales and reconstructs high-quality largescale WSIs. This is achieved using only an 8×8 vocabulary index for 256×256 images while maintaining minimal reconstruction loss. Experimental results demonstrate that our approach significantly enhances image quality, achieving an approximately 30% improvement in mean Fréchet inception distance (FID) compared to popular conditional GANs and diffusion models across various degradation scenarios in pathological images. Project website: https: //github.com/ziniBRC/CVARPath
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Autoregressive Image Generation without Vector QuantizationTianhong Li, Yonglong Tian, He Li, Mingyang Deng 等NeurIPS 2024 · 被引用 758 次
- DiffIR: Efficient Diffusion Model for Image RestorationBin Xia, Yulun Zhang, Shiyin Wang, Yitong Wang 等ICCV 2023 · 被引用 410 次
- Autoregressive Image Generation using Residual QuantizationDoyup Lee, Chiheon Kim, Saehoon Kim, Minsu Cho 等CVPR 2022 · 被引用 184 次
- AutoTimes: Autoregressive Time Series Forecasters via Large Language ModelsYong Liu, Guo Qin, Xiangdong Huang, Jianmin Wang 等NeurIPS 2024 · 被引用 138 次
相关 Paper
- PathVQ: Reforming Computational Pathology Foundation Model for Whole Slide Image Analysis via Vector QuantizationHonglin Li, Zhongyi Shui, Yunlong Zhang, Chenglu Zhu 等NeurIPS 2025 · 被引用 6 次
- Visual Autoregressive Modeling for Image Super-ResolutionYunpeng Qu, Kun Yuan, Jinhua Hao, Kai Zhao 等ICML 2025
- CARE: A Molecular-Guided Foundation Model with Adaptive Region Modeling for Whole Slide Image AnalysisDi Zhang, Zhangpeng Gong, Xiaobo Pang, Jiashuai Liu 等CVPR 2026 · 被引用 11 次
- TopoSlide: Topologically-Informed Histopathology Whole Slide Image Representation LearningShahira Abousamra, Asmita Sood, Sylvia PlevritisCVPR 2026
- GeneVAR: Causal MeanFlow for Autoregressive Gene-to-WSI Tile SynthesisJianwei Zhao, Fan Yang, Xin Li, Qiang Zhai 等CVPR 2026
