Lune

NeurIPS2025顶会

Can Dependencies Induced by LLM-Agent Workflows Be Trusted?

Yu Yao, Yiliao Song, Yian Xie, Mengdan Fan, Mingyu Guo, Tongliang Liu

2025年份
4被引次数

摘要

LLM-agent systems often decompose high-level objectives into subtask dependency graphs, assuming that each subtask’s output is reliable and conditionally independent of others given its parent responses. However, this assumption frequently breaks during execution, as ground-truth responses are inaccessible, leading to inter-agent misalignment —failures caused by inconsistencies and coordination breakdowns among agents [1]. To address this, we propose S EQ CV, a dynamic framework for reliable execution under violated conditional independence. S EQ CV executes subtasks sequentially, each conditioned on all prior verified responses, and performs consistency checks immediately after agents generate short token sequences. At each checkpoint, a token sequence is accepted only if it represents shared knowledge consistently supported across diverse LLM models; otherwise, it is discarded, triggering recursive subtask decomposition for finer-grained reasoning. Despite its sequential nature, S EQ CV avoids repeated corrections on the same misalignment and achieves higher effective throughput than parallel pipelines. Across multiple reasoning and coordination tasks, S EQ CV improves accuracy by up to 30% over existing LLM-agent systems. Code is available at github.com/tmllab/2025_NeurIPS_SeqCV .

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

lune papers fulltext b08354f7-d223-40fc-9105-b096e92e252d

它引用的顶会 Paper35

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖