Nested Browser-Use Learning for Agentic Information Seeking
Baixuan Li, Jialong Wu, Wenbiao Yin, Kuan Li, Zhongwang Zhang, Huifeng Yin, Zhengwei Tao, Liwen Zhang, Pengjun Xie, Jingren Zhou, Yong Jiang, Wentao Zhang, Zhiqiang Gao
Abstract
Information-seeking (IS) agents have achieved strong performance across a range of wide and deep search tasks, yet their tool use remains largely restricted to API-level snippet retrieval and URL-based page fetching, limiting access to the richer information available through real browsing. While full browser interaction could unlock deeper capabilities, its fine-grained control and verbose page content returns introduce substantial complexity for ReAct-style function-calling agents. To bridge this gap, we propose Nested Browser-Use Learning (NestBrowse), which introduces a minimal and complete browser-action framework that decouples interaction control from page exploration through a nested structure. This design simplifies agentic reasoning while enabling effective deep-web information acquisition. Empirical results on challenging deep IS benchmarks demonstrate that NestBrowse offers clear benefits in practice. Further in-depth analyses underscore its efficiency and flexibility.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- What Do Agents Learn from Trajectory-SFT: Semantics or Interfaces?Weizheng Gu, Chengze Li, Zhuohao Yu, Mengyuan Sun et al.ICML 2026
- RE-TRAC: REcursive TRAjectory Compression for Deep Search Agentsjialiang zhu, Gongrui Zhang, Xiaolong Ma, Lin Xu et al.ICML 2026
Builds on12
- Measuring Massive Multitask Language UnderstandingDan Hendrycks, Collin Burns, Steven Basart, Andy Zou et al.ICLR 2021 · 7,905 citations
- HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging FaceYongliang Shen, Kaitao Song, Xu Tan, Dongsheng Li et al.NeurIPS 2023 · 1,778 citations
- Gorilla: Large Language Model Connected with Massive APIsShishir G. Patil, Tianjun Zhang, Xin Wang, Joseph E. GonzalezNeurIPS 2024 · 1,715 citations
- AgentBench: Evaluating LLMs as AgentsXiao Liu, Hao Yu, Hanchen Zhang, Yifan Xu et al.ICLR 2024 · 748 citations
- GAIA: a benchmark for General AI AssistantsGrégoire Mialon, Clémentine Fourrier, Thomas Wolf, Yann LeCun et al.ICLR 2024 · 716 citations
Related papers
- WebClipper: Efficient Evolution of Web Agents with Graph-based Trajectory PruningJunjie Wang, Zequn Xie, Dan Yang, Jie Feng et al.ACL 2026
- Branch-and-Browse: Efficient and Controllable Web Exploration with Tree-Structured Reasoning and Action MemoryShiqi He, Yue Cui, Xinyu Ma, Yaliang Li et al.ACL 2026 · 5 citations
- FlowSearcher: Synthesizing Memory-Guided Agentic Workflows for Web Information SeekingKeyi Xiang, Zeyu Feng, Zhuoyi Lin, Yueming Lyu et al.ICLR 2026
- WALT: Web Agents that Learn ToolsViraj Prabhu, Yutong Dai, Matthew Fernandez, Krithika Ramakrishnan et al.ICLR 2026 · 13 citations
- Orca: Browsing at Scale Through User-Driven and AI-Facilitated Orchestration Across Malleable WebpagesPeiling Jiang, Haijun XiaCHI 2026 · 1 citation
