Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents
Mobile-Bench: оценочный бенчмарк для мобильных агентов на основе больших языковых моделей
2024-01-01
SCID: 54.1/6t5k5zv6
Discuss with AI
LLM-based mobile agentsMobile-Benchevaluation benchmarklarge language modelsmobile agent evaluation
Figures from the paper
Abstract (AI)
Shihan Deng, Weikai Xu, Hongda Sun, Wei Liu, Tao Tan, Jianfeng Liu, Ang Li, Jian Luan, Bin Wang, Rui Yan, Shuo Shang. Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2024.
Key Findings
1
Specific benchmark design, tasks, datasets, metrics, experimental results, comparisons, and limitations are not stated in the provided material.
2
The paper is identified as Mobile-Bench, an evaluation benchmark for LLM-based mobile agents, published at ACL 2024.
3
The provided text contains no abstract content, so substantive benchmark findings cannot be extracted reliably.
Research Object
LLM-based mobile agents
Research Subject
the evaluation performance and capabilities of LLM-based mobile agents across mobile tasks
Publication Details
Publication Date
2024-01-01
Journal
Publisher
ISSN
Cited by
17
Open access PDF
Access Type
Author Information
Download PDF
Subscribe to digest