AIR: Complex Instruction Generation via Automatic Iterative Refinement
Wei Liu, Yancheng He, Yu Li, Hui Huang, Chengwei Hu, Jiaheng Liu, Shilong Li, Wenbo Su, Bo Zheng
摘要
With the development of large language models, their ability to follow simple instructions has significantly improved.However, adhering to complex instructions remains a major challenge.Current approaches to generating complex instructions are often irrelevant to the current instruction requirements or suffer from limited scalability and diversity.Moreover, methods such as back-translation, while effective for simple instruction generation, fail to leverage the rich knowledge and formatting in human written documents.In this paper, we propose a novel Automatic Iterative Refinement (AIR) framework to generate complex instructions with constraints, which not only better reflects the requirements of real scenarios but also significantly enhances LLMs' ability to follow complex instructions.The AIR framework consists of two stages: 1) Generate an initial instruction from a document; 2) Iteratively refine instructions with LLM-as-judge guidance by comparing the model's output with the document to incorporate valuable constraints.Finally, we construct the AIR-10K dataset with 10K complex instructions and demonstrate that instructions generated with our approach significantly improve the model's ability to follow complex instructions, outperforming existing methods for instruction generation 1 .1 Codes and data are available at https://github.com/ WeiLiuAH/AIR-Automatic-Iterative-Refinement.Help me to write an advertisement line for laptop. Initial InstructionPower up your productivity and unleash creativity with our cutting-edge laptop-where performance meets portability!make the line with around 10 words.C1 Unleash your potential with speed, style, and innovation.Savior: Unleash the power of innovation in your hands.Savior: Unleash epic gaming performance with cutting-edge power and immersive visuals refer to the name of the laptop as savior.C2 C3 emphasize its gaming performance.C3
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Incentivizing Reasoning for Advanced Instruction-Following of Large Language ModelsYulei Qin, Gang Li, Zongyi Li, Zihan Xu 等NeurIPS 2025 · 被引用 17 次
- RECAST: Expanding the Boundaries of LLMs' Complex Instruction Following with Multi-Constraint DataZhengkang Guo, Wenhao Liu, Mingchen Xie, Jingwen Xu 等ICLR 2026 · 被引用 13 次
- Replay Failures as Successes: Sample-Efficient Reinforcement Learning for Instruction FollowingKongcheng Zhang, QI YAO, Shunyu Liu, Wenjian Zhang 等ICML 2026 · 被引用 4 次
- ELLMob: Event-Driven Human Mobility Generation with Self-Aligned LLM FrameworkYusong Wang, Chuang Yang, Jiawei Wang, Xiaohang Xu 等ICLR 2026 · 被引用 2 次
- DecIF: Improving Instruction-Following through DecompositionTingfeng Hui, Pengyu Zhu, Bowen Ping, Ling Tang 等ACL 2026
它引用的顶会 Paper13
- Measuring Massive Multitask Language UnderstandingDan Hendrycks, Collin Burns, Steven Basart, Andy Zou 等ICLR 2021 · 被引用 7,905 次
- Self-Refine: Iterative Refinement with Self-FeedbackAman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan 等NeurIPS 2023 · 被引用 4,972 次
- Self-Instruct: Aligning Language Models with Self-Generated InstructionsYizhong Wang, Yeganeh Kordi, Swaroop Mishra, Alisa Liu 等ACL 2023 · 被引用 540 次
- What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction TuningWei Liu, Weihao Zeng, Keqing He, Yong Jiang 等ICLR 2024 · 被引用 369 次
- Self-Alignment with Instruction BacktranslationXian Li, Ping Yu, Chunting Zhou, Timo Schick 等ICLR 2024 · 被引用 174 次
相关 Paper
- ToolACE-R: Model-aware Iterative Training and Adaptive Refinement for Tool learningXingshan Zeng, Weiwen Liu, Xu Huang, Zezhong Wang 等AAAI 2026 · 被引用 3 次
- From Selection to Refinement: Iterative Optimization for Instruction DataHang Hu, Ziyan Liu, Rujie Wen, Ruihui Hou 等ACL 2026
- Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis RefinementLinlu Qiu, Liwei Jiang, Ximing Lu, Melanie Sclar 等ICLR 2024 · 被引用 114 次
- A Stitch in Time Saves Nine: Proactive Self-Refinement for Language ModelsJinyi Han, Xinyi Wang, Haiquan Zhao, tingyun li 等ICLR 2026 · 被引用 2 次
- What Does LLM Refinement Actually Improve? A Systematic Study on Document-Level Literary TranslationShaomu Tan, Dawei Zhu, Ke Tran, Michael J. Denkowski 等ACL 2026 · 被引用 1 次
