Multi-Label Few-Shot ICD Coding as Autoregressive Generation with Prompt
Zhichao Yang, Sunjae Kwon, Zonghai Yao, Hong Yu
摘要
Automatic International Classification of Diseases (ICD) coding aims to assign multiple ICD codes to a medical note with an average of 3,000+ tokens. This task is challenging due to the high-dimensional space of multi-label assignment (155,000+ ICD code candidates) and the long-tail challenge - Many ICD codes are infrequently assigned yet infrequent ICD codes are important clinically. This study addresses the long-tail challenge by transforming this multi-label classification task into an autoregressive generation task. Specifically, we first introduce a novel pretraining objective to generate free text diagnoses and procedures using the SOAP structure, the medical logic physicians use for note documentation. Second, instead of directly predicting the high dimensional space of ICD codes, our model generates the lower dimension of text descriptions, which then infers ICD codes. Third, we designed a novel prompt template for multi-label classification. We evaluate our Generation with Prompt (GPsoap) model with the benchmark of all code assignment (MIMIC-III-full) and few shot ICD code assignment evaluation benchmark (MIMIC-III-few). Experiments on MIMIC-III-few show that our model performs with a marco F130.2, which substantially outperforms the previous MIMIC-III-full SOTA model (marco F1 4.3) and the model specifically designed for few/zero shot setting (marco F1 18.7). Finally, we design a novel ensemble learner, a cross-attention reranker with prompts, to integrate previous SOTA and our best few-shot coding predictions. Experiments on MIMIC-III-full show that our ensemble learner substantially improves both macro and micro F1, from 10.4 to 14.6 and from 58.2 to 59.1, respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper10
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained TransformersWenhui Wang, Furu Wei, Li Dong, Hangbo Bao 等NeurIPS 2020 · 被引用 2,727 次
- PEGASUS: Pre-training with Extracted Gap-sentences for Abstractive SummarizationJingqing Zhang, Yao Zhao, Mohammad Saleh, Peter J. LiuICML 2020 · 被引用 2,453 次
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- HyperCore: Hyperbolic and Co-graph Representation for Automatic ICD CodingPengfei Cao, Yubo Chen, Kang Liu, Jun Zhao 等ACL 2020 · 被引用 104 次
相关 Paper
- Automatic ICD Coding via Interactive Shared Representation Networks with Self-distillation MechanismTong Zhou, Pengfei Cao, Yubo Chen, Kang Liu 等ACL 2021
- ICD Coding from Clinical Text Using Multi-Filter Residual Convolutional Neural NetworkFei Li, Hong YuAAAI 2020 · 被引用 201 次
- Effective Convolutional Attention Network for Multi-label Clinical Document ClassificationYang Liu, Hua Cheng, Russell Klopfer, Matthew R. Gormley 等EMNLP 2021 · 被引用 51 次
- Less is More: Explainable and Efficient ICD Code Prediction with Clinical EntitiesJames C. Douglas, Yidong Gan, Ben Hachey, Jonathan K. KummerfeldACL 2025
- Coding Electronic Health Records with Adversarial Reinforcement Path GenerationShanshan Wang, Pengjie Ren, Zhumin Chen, Zhaochun Ren 等SIGIR 2020 · 被引用 19 次
