Beyond Code Generation: An Observational Study of ChatGPT Usage in Software Engineering Practice
Ranim Khojah, Mazen Mohamad, Philipp Leitner, Francisco Gomes de Oliveira Neto
摘要
Large Language Models (LLMs) are frequently discussed in academia and the general public as support tools for virtually any use case that relies on the production of text, including software engineering. Currently, there is much debate, but little empirical evidence, regarding the practical usefulness of LLM-based tools such as ChatGPT for engineers in industry. We conduct an observational study of 24 professional software engineers who have been using ChatGPT over a period of one week in their jobs, and qualitatively analyse their dialogues with the chatbot as well as their overall experience (as captured by an exit survey). We nd that rather than expecting ChatGPT to generate ready-to-use software artifacts (e.g., code), practitioners more often use ChatGPT to receive guidance on how to solve their tasks or learn about a topic in more abstract terms. We also propose a theoretical framework for how the (i) purpose of the interaction, (ii) internal factors (e.g., the user's personality), and (iii) external factors (e.g., company policy) together shape the experience (in terms of perceived usefulness and trust). We envision that our framework can be used by future research to further the academic discussion on LLM usage by software engineering practitioners, and to serve as a reference point for the design of future empirical LLM research in this domain. CCS Concepts: • Software and its engineering; • Human-centered computing → Natural language interfaces;
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Assistance or Disruption? Exploring and Evaluating the Design and Trade-offs of Proactive AI Programming SupportKevin Pu, Daniel Lazaro, Ian Arawjo, Haijun Xia 等CHI 2025 · 被引用 31 次
- "Create a Fear of Missing Out" - ChatGPT Implements Unsolicited Deceptive Designs in Generated Websites Without WarningVeronika Krauß, Mark McGill, Thomas Kosch, Yolanda Maira Thiel 等CHI 2025 · 被引用 17 次
- "Maybe We Need Some More Examples:" Individual and Team Drivers of Developer GenAI Tool UseCourtney Miller, Rudrajit Choudhuri, Mara Ulloa, Sankeerti Haniyur 等ICSE 2026 · 被引用 1 次
- Beyond the Desk: Barriers and Future Opportunities for AI to Assist Scientists in Embodied Physical TasksIrene Hou, Alexander Qin, Lauren Cheng, Philip J. GuoCHI 2026 · 被引用 1 次
- Toward Systematic Counterfactual Fairness Evaluation of Large Language Models: The CAFFE FrameworkAlessandra Parziale, Gianmario Voria, Valeria Pontillo, Gemma Catolino 等ICSE 2026
它引用的顶会 Paper7
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Grounded Copilot: How Programmers Interact with Code-Generating ModelsShraddha Barke, Michael B. James, Nadia PolikarpovaOOPSLA 2023 · 被引用 408 次
- CodaMosa: Escaping Coverage Plateaus in Test Generation with Pre-trained Large Language ModelsCaroline Lemieux, Jeevana Priya Inala, Shuvendu K. Lahiri, Siddhartha SenICSE 2023 · 被引用 221 次
- On the Robustness of Code Generation Techniques: An Empirical Study on GitHub CopilotAntonio Mastropaolo, Luca Pascarella, Emanuela Guglielmi, Matteo Ciniselli 等ICSE 2023 · 被引用 124 次
- An empirical study of bots in software development: characteristics and challenges from a practitioner's perspectiveLinda Erlenhov, Francisco Gomes de Oliveira Neto, Philipp LeitnerFSE 2020 · 被引用 47 次
相关 Paper
- Rocks Coding, Not Development: A Human-Centric, Experimental Evaluation of LLM-Supported SE TasksWei Wang, Huilong Ning, Gaowei Zhang, Libo Liu 等FSE 2024 · 被引用 17 次
- "If the Machine Is As Good As Me, Then What Use Am I?" - How the Use of ChatGPT Changes Young Professionals' Perception of Productivity and AccomplishmentCharlotte Kobiella, Yarhy Said Flores López, Franz Waltenberger, Fiona Draxler 等CHI 2024 · 被引用 54 次
- Evaluating Large Language Models on Academic Literature Understanding and Review: An Empirical Study among Early-stage ScholarsJiyao Wang, Haolong Hu, Zuyuan Wang, Song Yan 等CHI 2024 · 被引用 21 次
- Code Red! On the Harmfulness of Applying Off-the-Shelf Large Language Models to Programming TasksAli Al-Kaswan, Sebastian Deatc, Begüm Koç, Arie van Deursen 等FSE 2025 · 被引用 1 次
- Evaluating and Improving ChatGPT for Unit Test GenerationZhiqiang Yuan, Mingwei Liu, Shiji Ding, Kaixin Wang 等FSE 2024 · 被引用 89 次
