ChatBR: Automated assessment and improvement of bug report quality using ChatGPT
Lili Bo, Wangjie Ji, Xiaobing Sun, Ting Zhang, Xiaoxue Wu, Ying Wei
摘要
Bug reports, containing crucial information such as the Observed Behavior (OB), the Expected Behavior (EB), and the Steps to Reproduce (S2R), can help developers localize and fix bugs efficiently. However, due to the increasing complexity of some bugs and the limited experience of some reporters, many bug reports miss this crucial information. Although machine learning (ML)-based and information retrieval (IR)-based approaches have been proposed to detect and supplement the missing information in bug reports, the performance of these approaches depends heavily on the size and quality of bug report datasets.
In this paper, we present ChatBR, an approach for automated assessment and improvement of bug report quality using ChatGPT. First, we fine-tune a BERT model using manually annotated bug reports to create a sentence-level multi-label classifier to assess the quality of bug reports by detecting the presence of OB, EB, and S2R. Second, we use ChatGPT in a zero-shot setup to generate the missing information (OB, EB, and S2R) to improve the quality of bug reports. Finally, the output of ChatGPT is fed back into the classifier for verification until ChatGPT generates the missing information. Experimental results demonstrate ChatBR's superiority in both detecting and generating missing information in bug reports. For detection, ChatBR surpasses the state-of-the-art method, improving precision by 25.38% to 29.20%. In generating missing information, ChatBR achieves an average semantic similarity of 77.62% between generated and original content across six diverse projects. Furthermore, ChatBR can generate more than 99.9% of
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper4
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Large Language Models are Zero-Shot ReasonersTakeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo 等NeurIPS 2022 · 被引用 8,168 次
- Prompting Is All You Need: Automated Android Bug Replay with Large Language ModelsSidong Feng, Chunyang ChenICSE 2024 · 被引用 143 次
- Stay Professional and Efficient: Automatically Generate Titles for Your Bug ReportsSongqiang Chen, Xiaoyuan Xie, Bangguo Yin, Yuanxiang Ji 等ASE 2020 · 被引用 25 次
相关 Paper
- Toward interactive bug reporting for (android app) end-usersYang Song, Junayed Mahmud, Ying Zhou, Oscar Chaparro 等FSE 2022 · 被引用 29 次
- Automated Program Repair via Conversation: Fixing 162 out of 337 Bugs for $0.42 Each using ChatGPTChunqiu Steven Xia, Lingming ZhangISSTA 2024 · 被引用 105 次
- BugListener: Identifying and Synthesizing Bug Reports from Collaborative Live ChatsLin Shi, Fangwen Mu, Yumin Zhang, Ye Yang 等ICSE 2022 · 被引用 6 次
- Feedback-Driven Automated Whole Bug Report Reproduction for Android AppsDingbang Wang, Yu Zhao, Sidong Feng, Zhaoxu Zhang 等ISSTA 2024 · 被引用 16 次
- Automatically Reproducing Android Bug Reports using Natural Language Processing and Reinforcement LearningZhaoxu Zhang, Robert Winn, Yu Zhao, Tingting Yu 等ISSTA 2023 · 被引用 14 次
