Evaluating Large Language Models on Academic Literature Understanding and Review: An Empirical Study among Early-stage Scholars
Jiyao Wang, Haolong Hu, Zuyuan Wang, Song Yan, Youyu Sheng, Dengbo He
摘要
The rapid advancement of large language models (LLMs) such as ChatGPT makes LLM-based academic tools possible. However, little research has empirically evaluated how scholars perform different types of academic tasks with LLMs. Through an empirical study followed by a semi-structured interview, we assessed 48 early-stage scholars’ performance in conducting core academic activities (i.e., paper reading and literature reviews) under different levels of time pressure. Before conducting the tasks, participants received different training programs regarding the limitations and capabilities of the LLMs. After completing the tasks, participants completed an interview. Quantitative data regarding the influence of time pressure, task type, and training program on participants’ performance in academic tasks was analyzed. Semi-structured interviews provided additional information on the influential factors of task performance, participants’ perceptions of LLMs, and concerns about integrating LLMs into academic workflows. The findings can guide more appropriate usage and design of LLM-based tools in assisting academic work.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper8
- Understanding the LLM-ification of CHI: Unpacking the Impact of LLMs at CHI through a Systematic Literature ReviewRock Yuren Pang, Hope Schroeder, Kynnedy Simone Smith, Solon Barocas 等CHI 2025 · 被引用 51 次
- How CO2STLY Is CHI? The Carbon Footprint of Generative AI in HCI Research and What We Should Do About ItNanna Inie, Jeanette Falk, Raghavendra SelvanCHI 2025 · 被引用 33 次
- Co-Writing with AI, on Human Terms: Aligning Research with User Demands Across the Writing ProcessMohi Reza, Jeb Thomas-Mitchell, Peter Dushniku, Nathan Laundry 等CSCW 2025 · 被引用 29 次
- Small, Medium, Large? A Meta-Study of Effect Sizes at CHI to Aid Interpretation of Effect Sizes and Power CalculationAnna-Marie Ortloff, Florin Martius, Mischa Meier, Theo Raimbault 等CHI 2025 · 被引用 17 次
- Large Language Models for Automated Literature Review: An Evaluation of Reference Generation, Abstract Writing, and Review CompositionXuemei Tang, Xufeng Duan, Zhenguang G. CaiEMNLP 2025 · 被引用 5 次
相关 Paper
- Understanding the Effect of Risk Perception on the Acceptance and Use of Large Language Models Among University StudentsMichael T. Rücker, Carolin Büchting, Thomas KoschCSCW 2025 · 被引用 4 次
- "If the Machine Is As Good As Me, Then What Use Am I?" - How the Use of ChatGPT Changes Young Professionals' Perception of Productivity and AccomplishmentCharlotte Kobiella, Yarhy Said Flores López, Franz Waltenberger, Fiona Draxler 等CHI 2024 · 被引用 54 次
- An Empirical Study to Understand How Students Use ChatGPT for Writing EssaysAndrew Jelson, Daniel Manesh, Alice Jang, Daniel Dunlap 等CHI 2026 · 被引用 3 次
- Understanding the Role of Large Language Models in Personalizing and Scaffolding Strategies to Combat Academic ProcrastinationAnanya Bhattacharjee, Yuchen Zeng, Sarah Yi Xu, Dana Kulzhabayeva 等CHI 2024 · 被引用 39 次
- Beyond Code Generation: An Observational Study of ChatGPT Usage in Software Engineering PracticeRanim Khojah, Mazen Mohamad, Philipp Leitner, Francisco Gomes de Oliveira NetoFSE 2024 · 被引用 56 次
