Are LLM Web Search Engines Sustainable? A Web-Measurement Study of Real-Time Fetching
Abdur-Rahman Ibrahim Sayyid-Ali, Daanish Uddin Khan, Naveed Anwar Bhatti
摘要
Large language model (LLM) based answer engines offer real-time, conversational answers but raise concerns about scalability and web sustainability. Unlike traditional search that serves results from cached indices, these systems fetch pages anew for each query, creating redundant network traffic. We present the first measurement-driven study of commercial LLM answer engines, combining automated client-side tracing and controlled server-side audits. Analyzing ChatGPT and Claude across 1,000 queries, we find that both operate as meta-search layers heavily reliant on existing indices, fetching top-ranked pages with minimal caching. A human-equivalent cost model shows their per-query network footprint far exceeds that of human searchers, varying sharply by architecture. These results reveal the infrastructural burden of real-time fetching and motivate cooperative efficiency measures like shared caches, transparent retrieval standards, and publisher controls such as llms.txt, to make AI-augmented search more sustainable.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- TRACE: Evaluating Execution Efficiency of LLM-Based Code TranslationZhihao Gong, Zeyu Sun, Dong Huang, Qingyuan Liang 等ACL 2026 · 被引用 5 次
- Cheaply Estimating Inference Efficiency Metrics for Autoregressive Transformer ModelsDeepak Narayanan, Keshav Santhanam, Peter Henderson, Rishi Bommasani 等NeurIPS 2023 · 被引用 14 次
- Unveiling Environmental Impacts of Large Language Model Serving: A Functional Unit ViewYanran Wu, Inez Hua, Yi DingACL 2025 · 被引用 15 次
- IQA-EVAL: Automatic Evaluation of Human-Model Interactive Question AnsweringRuosen Li, Ruochen Li, Barry Wang, Xinya DuNeurIPS 2024 · 被引用 26 次
- TokenPowerBench: Benchmarking the Power Consumption of LLM InferenceChenxu Niu, Wei Zhang, Jie Li, Yongjian Zhao 等AAAI 2026 · 被引用 12 次
