LLM-Based Agents for Tool Learning: A Survey

Агенты на основе больших языковых моделей для обучения работе с инструментами: обзор
Weikai Xu, Chengrui Huang, Shen Gao, Shuo Shang
2025-06-26

LLM-based agentsmultimodal toolstool learningtool planningtool retrieval
Abstract Human beings capable of making and using tools can accomplish tasks far beyond their innate abilities, and this paradigm of integration with tools may not be limited to humans themselves. Recently, the large language model (LLM) has demonstrated immense potential across various fields with its unique planning and reasoning abilities. However, there are still many challenges beyond its capabilities due to deficiencies in its training data and inherent illusions. Thus, integrating LLMs and tools into tool learning agents has become a new emerging research direction. To this end, we present a systematic investigation and comprehensive review of tool-learning agents in this paper. We start by introducing the definition of the tool learning task for Agents and then illustrating the typical architecture of the tool-learning models. Since these tools are all defined by users, LLM does not know what tools there are and what their functions are. Thus, LLMs should first find appropriate tools and split the tool retrieval methods into two categories: training-based and non-training-based. To accurately complete the user task, it is important to decompose the task into several sub-tasks and execute them in the correct order. Following that, we introduce the tool planning methods and organize these works by whether they rely on the model’s inherent reasoning capabilities for planning or utilize external reasoning tools. Due to the rapid development of this field, we also introduce an emerging frontier direction: using multimodal tools for LLM. In addition, we compile current open-source benchmarks and evaluation metrics, focusing on their scale, composition, calculation methods, and assessment dimensions. Next, we introduce several application scenarios for the LLM-based tool learning methods. Finally, we discuss the safety and ethical issues involved in tool learning.
1
It formalizes the tool-learning task and organizes typical agent architectures for integrating LLMs with user-defined tools.
2
The paper provides a systematic investigation and comprehensive review of LLM-based tool-learning agents.
3
The survey covers multimodal tool use, open-source benchmarks, evaluation metrics, application scenarios, and safety and ethical issues.
4
Tool planning methods are organized by reliance on inherent model reasoning versus external reasoning tools for task decomposition and ordered execution.
5
Tool retrieval methods are categorized into training-based and non-training-based approaches because LLMs initially lack knowledge of available tools and their functions.

LLM-based tool-learning agents

Their architectures, tool retrieval and planning methods, multimodal tool use, benchmarks, evaluation metrics, applications, and safety and ethical issues

Publication Details
Publication Date
2025-06-26
Journal
Publisher
ISSN
Cited by
45
Access Type
Author Information
Authors
Weikai Xu
Chengrui Huang
Shen Gao
Shuo Shang
Explore further
Open the scid.ai AI chat with a ready-made request: it will find papers on a similar topic and help build a literature review.
Find similar papers in the chat
Make a presentation
100%