From language to action: a review of large language models as autonomous agents and tool users

От языка к действию: обзор больших языковых моделей как автономных агентов и пользователей инструментов
Sadia Sultana Chowa, Riasad Alvi, S M Asif Ur Rahman, Md Abdur Rahman, Mohaimenul Azam Khan Raiaan, Md Rafiqul Islam, Mukhtar Hussain, Sami Azam
2026-01-06

autonomous agentslarge language modelsmulti-agent systemstool useverifiable reasoning
Abstract The pursuit of human-level artificial intelligence (AI) has significantly advanced the development of autonomous agents and Large Language Models (LLMs). LLMs are now widely utilized as decision-making agents for their ability to interpret instructions, manage sequential tasks, and adapt through feedback. This review examines recent developments in employing LLMs as autonomous agents and tool users and comprises seven research questions. We only used the papers published between 2023 and 2025 in conferences of the A* and A-ranked and Q1 journals. A structured analysis of the LLM agents’ architectural design principles, dividing their applications into single-agent and multi-agent systems, and strategies for integrating external tools is presented. In addition, the cognitive mechanisms of LLMs, including reasoning, planning, and memory, and the impact of prompting methods and fine-tuning procedures on agent performance are also investigated. Furthermore, we have evaluated current benchmarks and assessment protocols and provided an analysis of 68 publicly available datasets to assess the performance of LLM-based agents in various tasks. In conducting this review, we have identified critical findings on verifiable reasoning of LLMs, the capacity for self-improvement, and the personalization of LLM-based agents. Finally, we have discussed ten future research directions to overcome these gaps.
1
It assesses current benchmarks and protocols alongside 68 publicly available datasets covering diverse LLM-agent tasks.
2
It organizes agent architectures into single-agent and multi-agent systems and examines strategies for integrating external tools.
3
The analysis identifies major gaps in verifiable reasoning, self-improvement, and personalization, and proposes ten future research directions.
4
The review analyzes 2023–2025 research on LLMs as autonomous agents and tool users, restricted to A*/A conferences and Q1 journals.
5
The review evaluates LLM-agent reasoning, planning, memory, prompting, and fine-tuning as factors influencing agent performance.

Large language model-based autonomous agents and tool-using systems

Architectures, cognitive mechanisms, tool-integration strategies, performance, and evaluation of LLM-based agents

Publication Details
Publication Date
2026-01-06
Journal
Publisher
ISSN
Access Type
Author Information
Authors
Sadia Sultana Chowa
Riasad Alvi
S M Asif Ur Rahman
Md Abdur Rahman
Mohaimenul Azam Khan Raiaan
Md Rafiqul Islam
Mukhtar Hussain
Sami Azam
Explore further
Open the scid.ai AI chat with a ready-made request: it will find papers on a similar topic and help build a literature review.
Find similar papers in the chat →
Make a presentation
100%