Multiple Instance Captioning: Learning Representations From Histopathology Textbooks and Articles
Jevgenij Gamper, Nasir M. Rajpoot
2021Year
11Top-tier citations
Abstract
Figure 1: Four samples from ARCH, a multiple instance captioning computational pathology dataset. Samples on the left and right each consist of four image instances with a single caption; top-middle shows an image-caption pair while bottom-middle contains two image instances with a single caption. Labeled in color are examples of common tasks within computational pathology: diagnostic (orange); detection & classification (cyan); descriptive (yellow); special cell detection (red).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c26996e9-91ee-45c0-8178-1be0f30d03a1Cited by top-tier papers11
- Multimodal Co-Attention Transformer for Survival Prediction in Gigapixel Whole Slide ImagesRichard J. Chen, Ming Y. Lu, Wei-Hung Weng, Tiffany Y. Chen et al.ICCV 2021 · 369 citations
- Node-aligned Graph Convolutional Network for Whole-slide Image Representation and ClassificationYonghang Guan, Jun Zhang, Kuan Tian, Sen Yang et al.CVPR 2022 · 65 citations
- ViLa-MIL: Dual-scale Vision-Language Multiple Instance Learning for Whole Slide Image ClassificationJiangbo Shi, Chen Li, Tieliang Gong, Yefeng Zheng et al.CVPR 2024 · 38 citations
- Patho-R1: A Multimodal Reinforcement Learning-Based Pathology Expert ReasonerWenchuan Zhang, Penghao Zhang, Jingru Guo, Tao Cheng et al.AAAI 2026 · 17 citations
- UniCell: Universal Cell Nucleus Classification via Prompt LearningJunjia Huang, Haofeng Li, Xiang Wan, Guanbin LiAAAI 2024 · 4 citations
Builds on5
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 2,416 citations
- Scaling and Benchmarking Self-Supervised Visual Representation LearningPriya Goyal, Dhruv Mahajan, Abhinav Gupta, Ishan MisraICCV 2019 · 429 citations
- VirTex: Learning Visual Representations From Textual AnnotationsKaran Desai, Justin JohnsonCVPR 2021
Related papers
- OCELOT: Overlapped Cell on Tissue Dataset for HistopathologyJeongun Ryu, Aaron Valero Puche, Jaewoong Shin, Seonwook Park et al.CVPR 2023
- Visual Language Pretrained Multiple Instance Zero-Shot Transfer for Histopathology ImagesMing Y. Lu, Bowen Chen, Andrew Zhang, Drew F. K. Williamson et al.CVPR 2023
- Do Multiple Instance Learning Models Transfer?Daniel Shao, Richard J. Chen, Andrew H. Song, Joel Runevic et al.ICML 2025
- CPath-Omni: A Unified Multimodal Foundation Model for Patch and Whole Slide Image Analysis in Computational PathologyYuxuan Sun, Yixuan Si, Chenglu Zhu, Xuan Gong et al.CVPR 2025
- CAMEL: A Weakly Supervised Learning Framework for Histopathology Image SegmentationGang Xu, Zhigang Song, Zhuo Sun, Calvin Ku et al.ICCV 2019 · 187 citations
