Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
Lakshmi Nair, Ian Trase, J. Mark Kim
摘要
We present a novel reasoning approach called Flow-of-Options (FoO), designed to address intrinsic biases in Large Language Models (LLMs). Flow-of-Options enables LLMs to systematically explore a diverse range of possibilities in their reasoning, as demonstrated by an FoO-based agentic framework developed for autonomously solving Machine Learning (ML) tasks. FoO enforces diversity in LLM solutions through compressed and interpretable task representations, resulting in improvements of 38.2% -69.2% on standard data science tasks, and 37.4% -47.9% on therapeutic chemistry tasks, as compared to state-of-the-art baselines. With an overall operation cost under $1 per task, our framework is well-suited for cost-sensitive applications. Going beyond tabular classification and regression, we show the broader applicability of our FoO-based agentic system to tasks such as reinforcement learning and image generation. Our code is open-sourced at: https: //github.com/flagshippioneering/ Flow-of-Options.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Graph of Thoughts: Solving Elaborate Problems with Large Language ModelsMaciej Besta, Nils Blach, Ales Kubicek, Robert Gerstenberger 等AAAI 2024 · 被引用 1,292 次
- Grounding Large Language Models in Interactive Environments with Online Reinforcement LearningThomas Carta, Clément Romac, Thomas Wolf, Sylvain Lamprier 等ICML 2023 · 被引用 258 次
- Chain of Code: Reasoning with a Language Model-Augmented Code EmulatorChengshu Li, Jacky Liang, Andy Zeng, Xinyun Chen 等ICML 2024 · 被引用 155 次
- DS-Agent: Automated Data Science by Empowering Large Language Models with Case-Based ReasoningSiyuan Guo, Cheng Deng, Ying Wen, Hechang Chen 等ICML 2024 · 被引用 107 次
- Language Models as Compilers: Simulating Pseudocode Execution Improves Algorithmic Reasoning in Language ModelsHyungjoo Chae, Yeonghyeon Kim, Seungone Kim, Kai Tzu-iunn Ong 等EMNLP 2024
相关 Paper
- Fleet of Agents: Coordinated Problem Solving with Large Language ModelsLars Henning Klein, Nearchos Potamitis, Roland C. Aydin, Robert West 等ICML 2025
- FlowRL: Matching Reward Distributions for LLM ReasoningXuekai Zhu, Daixuan Cheng, Dinghuai Zhang, Hengli Li 等ICLR 2026 · 被引用 41 次
- Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal ExamplesFangxu Yu, Lai Jiang, Haoqiang Kang, Shibo Hao 等ICML 2025
- AFlow: Automating Agentic Workflow GenerationJiayi Zhang, Jinyu Xiang, Zhaoyang Yu, Fengwei Teng 等ICLR 2025
- Count Counts: Motivating Exploration in LLM Reasoning with Count-based Intrinsic RewardsXuan Zhang, Ruixiao Li, Zhijian Zhou, Long Li 等ICLR 2026 · 被引用 11 次
