SEE: Structure-aware Exploring and Exploiting for Long-horizon GUI Agent Trajectory Synthesis
Zhuohang Fan , Beichen Zhang , Yuanfa Li , Changqiao Wu , Wei Liu , Jian Luan , Weigang Zhang
- 🏛 Institutions
- Unknown
- 📅 Date
- July 20, 2026
- 📑 Publisher
- arXiv
- 💻 Env
- Mobile
- 🔑 Keywords
TLDR
SEE builds UI transition graphs through structured exploration, then synthesizes diverse long-horizon mobile trajectories with graph-based planning and controlled sampling. The method targets coverage of rare transitions while avoiding spurious interaction cycles.
Related papers (24)
- WebWorld: A Large-Scale World Model for Web Agent TrainingFebruary 16, 2026 · arXiv
- HATS: Hardness-Aware Trajectory Synthesis for GUI AgentsMarch 12, 2026 · CVPR 2026
- UI-Oceanus: Scaling GUI Agents with Synthetic Environmental DynamicsFebruary 11, 2026 · arXiv
- Learning with Challenges: Adaptive Difficulty-Aware Data Generation for Mobile GUI Agent TrainingJanuary 30, 2026 · arXiv
- Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent PretrainingMay 14, 2026 · arXiv
- AutoWebWorld: Synthesizing Infinite Verifiable Web Environments via Finite State MachinesFebruary 15, 2026 · arXiv
- Scaling Web Agent Training through Automatic Data Generation and Fine-grained EvaluationFebruary 13, 2026 · COLM 2025
- ANCHOR: Branch-Point Data Generation for GUI AgentsFebruary 6, 2026 · arXiv
- InfiniteWeb: Scalable Web Environment Synthesis for GUI Agent TrainingJanuary 7, 2026 · arXiv
- Explorer: Scaling Exploration-driven Web Trajectory Synthesis for Multimodal Web AgentsJuly 2025 · Findings of ACL 2025
- OS-Genesis: Automating GUI Agent Trajectory Construction via Reverse Task SynthesisDecember 27, 2024 · ACL 2025
- BlueLM-GUI Technical Report: A Real-Device-Centric Flywheel for Self-Improving Mobile GUI AgentsSeptember 11, 2026 · arXiv
- JarvisGUI: Towards Cross-Device GUI Agents with Dynamic Task CompositionSeptember 9, 2026 · arXiv
- APPSim-Bench: Bridging Real-world Apps and Reproducible Evaluation for Mobile GUI AgentsSeptember 7, 2026 · arXiv
- Improving Proficiency and Efficiency of Android GUI Agents via Self-Generating Tool ActionsSeptember 6, 2026 · arXiv
- ElderBench: Benchmarking Autonomous Mobile Agents for Older AdultsSeptember 4, 2026 · arXiv
- WiP: Characterizing and Defending Against Mobile-Agent-Driven MFA AutomationSeptember 2, 2026 · arXiv
- GUI-CC: Benchmarking Contextual Consistency of GUI World Models as Agent EnvironmentsAugust 30, 2026 · arXiv
- ActReal: System-Level Mobile Agents Challenge Mobile Automation DetectionAugust 30, 2026 · arXiv
- WM-R1: Training GUI Agents to Reason and leverage World Models with Reinforcement LearningAugust 27, 2026 · arXiv
- Are Android GUI Agents Robust Against Runtime Anomalies? AnTrap: Evaluating Agents in Dynamic Adversarial EnvironmentsAugust 25, 2026 · arXiv
- ADeptS-Bench: Measuring the Trustworthiness of Computer Use Agents Across DevicesAugust 25, 2026 · arXiv
- GSAR: Goal-State-Anchor Rewards for Mobile GUI Agents with Self-Evolving Data SynthesisAugust 24, 2026 · arXiv
- Lexical Coupling in GUI Element Grounding: Sentence Embeddings Track Labels across Mobile and WebAugust 22, 2026 · arXiv