Multi-Label Few-Shot ICD Coding as Autoregressive Generation with Prompt
Zhichao Yang, Sunjae Kwon, Zonghai Yao, Hong Yu
Abstract
Automatic International Classification of Diseases (ICD) coding aims to assign multiple ICD codes to a medical note with an average of 3,000+ tokens. This task is challenging due to the high-dimensional space of multi-label assignment (155,000+ ICD code candidates) and the long-tail challenge - Many ICD codes are infrequently assigned yet infrequent ICD codes are important clinically. This study addresses the long-tail challenge by transforming this multi-label classification task into an autoregressive generation task. Specifically, we first introduce a novel pretraining objective to generate free text diagnoses and procedures using the SOAP structure, the medical logic physicians use for note documentation. Second, instead of directly predicting the high dimensional space of ICD codes, our model generates the lower dimension of text descriptions, which then infers ICD codes. Third, we designed a novel prompt template for multi-label classification. We evaluate our Generation with Prompt (GPsoap) model with the benchmark of all code assignment (MIMIC-III-full) and few shot ICD code assignment evaluation benchmark (MIMIC-III-few). Experiments on MIMIC-III-few show that our model performs with a marco F130.2, which substantially outperforms the previous MIMIC-III-full SOTA model (marco F1 4.3) and the model specifically designed for few/zero shot setting (marco F1 18.7). Finally, we design a novel ensemble learner, a cross-attention reranker with prompts, to integrate previous SOTA and our best few-shot coding predictions. Experiments on MIMIC-III-full show that our ensemble learner substantially improves both macro and micro F1, from 10.4 to 14.6 and from 58.2 to 59.1, respectively.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e9282fa2-5f97-48ee-96de-5676827bc6bfCited by top-tier papers1
Ask how each one uses itBuilds on10
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained TransformersWenhui Wang, Furu Wei, Li Dong, Hangbo Bao et al.NeurIPS 2020 · 2,727 citations
- PEGASUS: Pre-training with Extracted Gap-sentences for Abstractive SummarizationJingqing Zhang, Yao Zhao, Mohammad Saleh, Peter J. LiuICML 2020 · 2,453 citations
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- HyperCore: Hyperbolic and Co-graph Representation for Automatic ICD CodingPengfei Cao, Yubo Chen, Kang Liu, Jun Zhao et al.ACL 2020 · 104 citations
Related papers
- Automatic ICD Coding via Interactive Shared Representation Networks with Self-distillation MechanismTong Zhou, Pengfei Cao, Yubo Chen, Kang Liu et al.ACL 2021
- ICD Coding from Clinical Text Using Multi-Filter Residual Convolutional Neural NetworkFei Li, Hong YuAAAI 2020 · 201 citations
- Effective Convolutional Attention Network for Multi-label Clinical Document ClassificationYang Liu, Hua Cheng, Russell Klopfer, Matthew R. Gormley et al.EMNLP 2021 · 51 citations
- Less is More: Explainable and Efficient ICD Code Prediction with Clinical EntitiesJames C. Douglas, Yidong Gan, Ben Hachey, Jonathan K. KummerfeldACL 2025
- Coding Electronic Health Records with Adversarial Reinforcement Path GenerationShanshan Wang, Pengjie Ren, Zhumin Chen, Zhaochun Ren et al.SIGIR 2020 · 19 citations
