VADIS: A Visual Analytics Pipeline for Dynamic Document Representation and Information-Seeking
Rui Qiu, Yamei Tu, Po-Yin Yen, Han-Wei Shen
Abstract
In the biomedical domain, visualizing the document embeddings of an extensive corpus has been widely used in information-seeking tasks. However, three key challenges with existing visualizations make it difficult for clinicians to find information efficiently. First, the document embeddings used in these visualizations are generated statically by pretrained language models, which cannot adapt to the user's evolving interest. Second, existing document visualization techniques cannot effectively display how the documents are relevant to users' interest, making it difficult for users to identify the most pertinent information. Third, existing embedding generation and visualization processes suffer from a lack of interpretability, making it difficult to understand, trust and use the result for decision-making. In this paper, we present a novel visual analytics pipeline for user-driven document representation and iterative information seeking (VADIS). VADIS introduces a prompt-based attention model (PAM) that generates dynamic document embedding and document relevance adjusted to the user's query. To effectively visualize these two pieces of information, we design a new document map that leverages a circular grid layout to display documents based on both their relevance to the query and the semantic similarity. Additionally, to improve the interpretability, we introduce a corpus-level attention visualization method to improve the user's understanding of the model focus and to enable the users to identify potential oversight. This visualization, in turn, empowers users to refine, update and introduce new queries, thereby facilitating a dynamic and iterative information-seeking experience. We evaluated VADIS quantitatively and qualitatively on a real-world dataset of biomedical research papers to demonstrate its effectiveness.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 087c174e-0f12-4c24-8bba-803569d9dc9eBuilds on4
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Dense Passage Retrieval for Open-Domain Question AnsweringVladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis et al.EMNLP 2020 · 142 citations
- VITALITY: Promoting Serendipitous Discovery of Academic Literature with Transformers & Visual AnalyticsArpit Narechania, Alireza Karduni, Ryan Wesslen, Emily WallIEEE VIS 2021 · 31 citations
- Bursting Scientific Filter Bubbles: Boosting Innovation via Novel Author DiscoveryJason Portenoy, Marissa Radensky, Jevin D. West, Eric Horvitz et al.CHI 2022 · 28 citations
Related papers
- Modeling and Leveraging Analytic Focus During Exploratory Visual AnalysisZhilan Zhou, Ximing Wen, Yue Wang, David GotzCHI 2021 · 21 citations
- MeDKCoOp: Dual Knowledge-guided Graph Prompt Learning for Biomedical Vision-Language ModelsYijun Wang, Siying Wu, Lubin Gan, Zheyu Zhang et al.ACM MM 2025
- Seeing the Trees for the Forest: Rethinking Weakly-Supervised Medical Visual GroundingTa Duc Huy, Duy Anh Huynh, Yutong Xie, Yuankai Qi et al.ICCV 2025 · 2 citations
- An Image is Worth Multiple Words: Discovering Object Level Concepts using Multi-Concept Prompt LearningChen Jin, Ryutaro Tanno, Amrutha Saseendran, Tom Diethe et al.ICML 2024 · 15 citations
- MedCLIPSeg: Probabilistic Vision-Language Adaptation for Data-Efficient and Generalizable Medical Image SegmentationTaha Koleilat, Hojat Asgariandehkordi, Omid Nejatimanzari, Berardino Barile et al.CVPR 2026 · 4 citations
