GUI Agents Papers
Star · 902

SeekJudge: A Practical Reward Framework for Reinforcement Learning in Computer-Use Agents

Yang Wan , Zhenhao Zhang , Jierui Wang , Linchao Zhu

🏛 Institutions
Unknown
📅 Date
July 25, 2026
📑 Publisher
arXiv
💻 Env
Desktop
🔑 Keywords
TLDR

SeekJudge uses specialized Condense, Ground, Seek, and Analyze roles to judge whether long computer-use trajectories satisfy their instructions. Its distilled shared-backbone design supplies step-level reward signals intended to make model-based supervision practical for online RL.

Open paper Report issue
Related papers (24)