Improving Segmentation for Technical Support Problems
Kushal Chauhan, Abhirut Gupta
Abstract
Technical support problems are often long and complex. They typically contain user descriptions of the problem, the setup, and steps for attempted resolution. Often they also contain various non-natural language text elements like outputs of commands, snippets of code, error messages or stack traces. These elements contain potentially crucial information for problem resolution. However, they cannot be correctly parsed by tools designed for natural language. In this paper, we address the problem of segmentation for technical support questions. We formulate the problem as a sequence labelling task, and study the performance of state of the art approaches. We compare this against an intuitive contextual sentence-level classification baseline, and a state of the art supervised text-segmentation approach. We also introduce a novel component of combining contextual embeddings from multiple language models pre-trained on different data sources, which achieves a marked improvement over using embeddings from a single pre-trained language model. Finally, we also demonstrate the usefulness of such segmentation with improvements on the downstream task of answer retrieval.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Related papers
- Automatic sentence segmentation of clinical record narratives in real-world dataDongfang Xu, Davy Weissenbacher, Karen O'Connor, Siddharth Rawal et al.EMNLP 2024 · 1 citation
- A unified approach to sentence segmentation of punctuated text in many languagesRachel Wicks, Matt PostACL 2021
- LISA: Reasoning Segmentation via Large Language ModelXin Lai, Zhuotao Tian, Yukang Chen, Yanwei Li et al.CVPR 2024
- Character-level Representations Improve DRS-based Semantic Parsing Even in the Age of BERTRik van Noord, Antonio Toral, Johan BosEMNLP 2020 · 22 citations
- The Inductive Bias of In-Context Learning: Rethinking Pretraining Example DesignYoav Levine, Noam Wies, Daniel Jannai, Dan Navon et al.ICLR 2022 · 43 citations
