GUI Agents Papers
Star · 902

APPSim-Bench: Bridging Real-world Apps and Reproducible Evaluation for Mobile GUI Agents

Jintian Feng , Long Chen , Xiao Yu , Jiayi Dai , Chenglong Liu , Haoru Wang , Zizhen Xue , Yuxuan Shi , Ziyang Wang , Yichen Gong

🏛 Institutions
Central China Normal University , Agentic Labs, Acrab AI
📅 Date
September 7, 2026
📑 Publisher
arXiv
💻 Env
Mobile
🔑 Keywords
TLDR

APPSim-Bench provides 557 tasks across 17 high-frequency Chinese and English apps, rebuilt as controllable simulated apps that preserve interaction logic while removing stochasticity from ads, recommendations, and accounts, and evaluates 19 GUI agents on them.

Open paper Report issue
Related papers (24)