GUI Agents Papers
Star · 902

Beyond Success and Failure: Length-Aware Contrastive Learning for GUI Agents

Chengyang Gu , Le Zhang , Jingbo Zhou , Yize Chen , Yu Shi , Siqi Bao , Zheng-Fan Wu , Hua Wu , Hui Xiong

🏛 Institutions
The Hong Kong University of Science and Technology (Guangzhou) , Baidu Inc. , University of Alberta
📅 Date
August 22, 2026
📑 Publisher
arXiv
💻 Env
General GUI
🔑 Keywords
TLDR

This work identifies a reward-gradient misalignment in GRPO-style GUI-agent RL caused by trajectory length and proposes length-aware contrastive learning that credits partial progress instead of collapsing every trajectory to a binary success/failure signal.

Open paper Report issue
Related papers (24)