Lune

NeurIPS2025Top-tier venue

Can Dependencies Induced by LLM-Agent Workflows Be Trusted?

Yu Yao, Yiliao Song, Yian Xie, Mengdan Fan, Mingyu Guo, Tongliang Liu

2025Year
4Citations

Abstract

LLM-agent systems often decompose high-level objectives into subtask dependency graphs, assuming that each subtask’s output is reliable and conditionally independent of others given its parent responses. However, this assumption frequently breaks during execution, as ground-truth responses are inaccessible, leading to inter-agent misalignment —failures caused by inconsistencies and coordination breakdowns among agents [1]. To address this, we propose S EQ CV, a dynamic framework for reliable execution under violated conditional independence. S EQ CV executes subtasks sequentially, each conditioned on all prior verified responses, and performs consistency checks immediately after agents generate short token sequences. At each checkpoint, a token sequence is accepted only if it represents shared knowledge consistently supported across diverse LLM models; otherwise, it is discarded, triggering recursive subtask decomposition for finer-grained reasoning. Despite its sequential nature, S EQ CV avoids repeated corrections on the same misalignment and achieves higher effective throughput than parallel pipelines. Across multiple reasoning and coordination tasks, S EQ CV improves accuracy by up to 30% over existing LLM-agent systems. Code is available at github.com/tmllab/2025_NeurIPS_SeqCV .

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext b08354f7-d223-40fc-9105-b096e92e252d

Builds on35

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines