Lune

EMNLP2024顶会

Towards Tool Use Alignment of Large Language Models

Zhiyuan Chen, Shiqi Shen, Guangyao Shen, Gong Zhi, Xu Chen, Yankai Lin

2024年份
5被引次数
6顶会引用

摘要

Recently, tool use with LLMs has become one of the primary research topics as it can help LLM generate truthful and helpful responses.Existing studies on tool use with LLMs primarily focus on enhancing the tool-calling ability of LLMs.In practice, like chat assistants, LLMs are also required to align with human values in the context of tool use.Specifically, LLMs should refuse to answer unsafe tool use relevant instructions and insecure tool responses to ensure their reliability and harmlessness.At the same time, LLMs should demonstrate autonomy in tool use to reduce the costs associated with tool calling.To tackle this issue, we first introduce the principle that LLMs should follow in tool use scenarios: H2A.The goal of H2A is to align LLMs with helpfulness, harmlessness, and autonomy.In addition, we propose ToolAlign, a dataset comprising instruction-tuning data and preference data to align LLMs with the H2A principle for tool use.Based on ToolAlign, we develop LLMs by supervised fine-tuning and preference learning, and experimental results demonstrate that the LLMs exhibit remarkable toolcalling capabilities, while also refusing to engage with harmful content, and displaying a high degree of autonomy in tool utilization.The code and datasets are available at: https: //github.com/zhiyuanc2001/ToolAlign.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper6

问问它们各自怎么用它

它引用的顶会 Paper14

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖