Prioritize Crowdsourced Test Reports via Deep Screenshot Understanding
Shengcheng Yu, Chunrong Fang, Zhenfei Cao, Xu Wang, Tongyu Li, Zhenyu Chen
Abstract
Crowdsourced testing is increasingly dominant in mobile application (app) testing, but it is a great burden for app developers to inspect the incredible number of test reports. Many researches have been proposed to deal with test reports based only on texts or additionally simple image features. However, in mobile app testing, texts contained in test reports are condensed and the information is inadequate. Many screenshots are included as complements that contain much richer information beyond texts. This trend motivates us to prioritize crowdsourced test reports based on a deep screenshot understanding. In this paper, we present a novel crowdsourced test report prioritization approach, namely DeepPrior. We fifirstrst represent the crowdsourced test reports with a novelly introduced feature, namely DeepFeature, that includes all the widgets along with their texts, coordinates, types, and even intents based on the deep analysis of the app screenshots, and the textual descriptions in the crowdsourced test reports. DeepFeature includes the <i xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">Bug</i> <i xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">Feature</i> , which directly describes the bugs, and the <i xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">Context</i> <i xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">Feature</i> , which depicts the thorough context of the bug. The similarity of the DeepFeature is used to represent the test reports' similarity and prioritize the crowdsourced test reports. We formally define the similarity as DeepSimilarity. We also conduct an empirical experiment to evaluate the effectiveness of the proposed technique with a large dataset group. The results show that DeepPrior is promising, and it outperforms the state-of-the-art approach with less than half the overhead.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 23f91bbb-72cc-43a3-a64f-449c05105697Cited by top-tier papers2
- Practical Non-Intrusive GUI Exploration Testing with Visual-based Robotic ArmsShengcheng Yu, Chunrong Fang, Mingzhe Du, Yuchen Ling et al.ICSE 2024 · 9 citations
- Towards Automated Crowdsourced Testing via Personified-LLMShengcheng Yu, Yuchen Ling, Chunrong Fang, Zhenyu Chen et al.FSE 2026 · 1 citation
Related papers
- Semi-supervised Crowdsourced Test Report Clustering via Screenshot-Text Binding RulesShengcheng Yu, Chunrong Fang, Quanjun Zhang, Mingzhe Du et al.FSE 2024 · 4 citations
- Automatically Matching Bug Reports With Related App ReviewsMarlo Haering, Christoph Stanik, Walid MaalejICSE 2021 · 52 citations
- Multi-dimensional Assessment of Crowdsourced Testing Reports via LLMsYue Wang, Yuan Zhang, Shengcheng Yu, Zhenyu ChenASE 2025
- CertPri: Certifiable Prioritization for Deep Neural Networks via Movement Cost in Feature SpaceHaibin Zheng, Jinyin Chen, Haibo JinASE 2023 · 11 citations
- Owl Eyes: Spotting UI Display Issues via Visual UnderstandingZhe Liu, Chunyang Chen, Junjie Wang, Yuekai Huang et al.ASE 2020 · 79 citations
