Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Mobile-Bench: оценочный бенчмарк для мобильных агентов на основе больших языковых моделей
Shihan Deng, Weikai Xu, Hongda Sun, Wei Liu, Tao Tan, Liujianfeng Liujianfeng, Ang Li, Jian Luan, Bin Wang, Rui Yan, Shuo Shang
2024-01-01

LLM-based mobile agentsMobile-Benchevaluation benchmarklarge language modelsmobile agent evaluation
Shihan Deng, Weikai Xu, Hongda Sun, Wei Liu, Tao Tan, Jianfeng Liu, Ang Li, Jian Luan, Bin Wang, Rui Yan, Shuo Shang. Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2024.
1
Specific benchmark design, tasks, datasets, metrics, experimental results, comparisons, and limitations are not stated in the provided material.
2
The paper is identified as Mobile-Bench, an evaluation benchmark for LLM-based mobile agents, published at ACL 2024.
3
The provided text contains no abstract content, so substantive benchmark findings cannot be extracted reliably.

LLM-based mobile agents

the evaluation performance and capabilities of LLM-based mobile agents across mobile tasks

Publication Details
Publication Date
2024-01-01
Journal
Publisher
ISSN
Cited by
17
Access Type
Author Information
Authors
Shihan Deng
Weikai Xu
Hongda Sun
Wei Liu
Tao Tan
Liujianfeng Liujianfeng
Ang Li
Jian Luan
Bin Wang
Rui Yan
Shuo Shang
Explore further
Open the scid.ai AI chat with a ready-made request: it will find papers on a similar topic and help build a literature review.
Find similar papers in the chat
Make a presentation
100%