Semi-Supervised Variational Reasoning for Medical Dialogue Generation
Dongdong Li, Zhaochun Ren, Pengjie Ren, Zhumin Chen, Miao Fan, Jun Ma, Maarten de Rijke
Abstract
Medical dialogue generation aims to provide automatic and accurate responses to assist physicians to obtain diagnosis and treatment suggestions in an efficient manner. In medical dialogues, two key characteristics are relevant for response generation: patient states (such as symptoms, medication) and physician actions (such as diagnosis, treatments). In medical scenarios large-scale human annotations are usually not available, due to the high costs and privacy requirements. Hence, current approaches to medical dialogue generation typically do not explicitly account for patient states and physician actions, and focus on implicit representation instead. We propose an end-to-end variational reasoning approach to medical dialogue generation. To be able to deal with a limited amount of labeled data, we introduce both patient state and physician action as latent variables with categorical priors for explicit patient state tracking and physician policy learning, respectively. We propose a variational Bayesian generative approach to approximate posterior distributions over patient states and physician actions. We use an efficient stochastic gradient variational Bayes estimator to optimize the derived evidence lower bound, where a 2-stage collapsed inference method is proposed to reduce the bias during model training. A physician policy network composed of an action-classifier and two reasoning detectors is proposed for augmented reasoning ability. We conduct experiments on three datasets collected from medical platforms. Our experimental results show that the proposed method outperforms state-of-the-art baselines in terms of objective and subjective evaluation metrics. Our experiments also indicate that our proposed semi-supervised reasoning method achieves a comparable performance as state-of-the-art fully supervised learning baselines for physician policy learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- MediQ: Question-Asking LLMs and a Benchmark for Reliable Interactive Clinical ReasoningShuyue Stella Li, Vidhisha Balachandran, Shangbin Feng, Jonathan Ilgen et al.NeurIPS 2024 · 215 citations
- DrHouse: An LLM-empowered Diagnostic Reasoning System through Harnessing Outcomes from Sensor Data and Expert KnowledgeBufang Yang, Siyang Jiang, Lilin Xu, Kaiwei Liu et al.UbiComp 2025 · 60 citations
- Transfer Learning with Synthetic Corpora for Spatial Role Labeling and ReasoningRoshanak Mirzaee, Parisa KordjamshidiEMNLP 2022 · 10 citations
- Doctor-R1: Mastering Clinical Inquiry with Experiential Agentic Reinforcement LearningYunghwei Lai, Kaiming Liu, Ziyue Wang, Weizhi Ma et al.ICLR 2026 · 10 citations
- CDialog: A Multi-turn Covid-19 Conversation Dataset for Entity-Aware Dialog GenerationDeeksha Varshney, Aizan Zafar, Niranshu Kumar Behra, Asif EkbalEMNLP 2022 · 2 citations
Builds on13
- A Simple Language Model for Task-Oriented DialogueEhsan Hosseini-Asl, Bryan McCann, Chien-Sheng Wu, Semih Yavuz et al.NeurIPS 2020 · 590 citations
- Task-Oriented Dialog Systems That Consider Multiple Appropriate Responses under the Same ContextYichi Zhang, Zhijian Ou, Zhou YuAAAI 2020 · 198 citations
- Sequential Latent Knowledge Selection for Knowledge-Grounded DialogueByeongchang Kim, Jaewoo Ahn, Gunhee KimICLR 2020 · 179 citations
- Interactive Path Reasoning on Graph for Conversational RecommendationWenqiang Lei, Gangyi Zhang, Xiangnan He, Yisong Miao et al.KDD 2020 · 158 citations
- Generative Adversarial Regularized Mutual Information Policy Gradient Framework for Automatic DiagnosisYuan Xia, Jingbo Zhou, Zhenhui Shi, Chao Lu et al.AAAI 2020 · 85 citations
Related papers
- Variational Hierarchical Dialog Autoencoder for Dialog State Tracking Data AugmentationKang Min Yoo, Hanbit Lee, Franck Dernoncourt, Trung Bui et al.EMNLP 2020
- A Probabilistic End-To-End Task-Oriented Dialog Model with Latent Belief States towards Semi-Supervised LearningYichi Zhang, Zhijian Ou, Min Hu, Junlan FengEMNLP 2020 · 52 citations
- MedDialog: Large-scale Medical Dialogue DatasetsGuangtao Zeng, Wenmian Yang, Zeqian Ju, Yue Yang et al.EMNLP 2020 · 163 citations
- Open Domain Dialogue Generation with Latent ImagesZe Yang, Wei Wu, Huang Hu, Can Xu et al.AAAI 2021 · 30 citations
- Diverse and Faithful Knowledge-Grounded Dialogue Generation via Sequential Posterior InferenceYan Xu, Deqian Kong, Dehong Xu, Ziwei Ji et al.ICML 2023 · 9 citations
