Synthesizing Agentic Data for Web Agents with Progressive Difficulty Enhancement Mechanisms
Shrey Pandit , Xuan-Phi Nguyen , Yifei Ming , Austin Xu , Jiayu Wang , Caiming Xiong , Shafiq Joty
- 🏛 Institutions
- Salesforce AI Research , University of Wisconsin-Madison
- 📅 Date
- October 15, 2025
- 📑 Publisher
- arXiv
- 💻 Env
- 🔑 Keywords
TLDR
This paper synthesizes training data for deep-research web agents by progressively increasing question difficulty until a baseline agent fails, then using that agent again for validation and filtering. The resulting corpus is aimed at long-horizon online-tool use rather than browser-native GUI interaction, but it is relevant as adjacent training-data work for agent systems.
Related papers (6)
- BrowserForge: Scaling Web Episode via Parallel Browser SandboxesAugust 25, 2026 · arXiv
- GSAR: Goal-State-Anchor Rewards for Mobile GUI Agents with Self-Evolving Data SynthesisAugust 24, 2026 · arXiv
- Training Needs Trustworthy Worlds: Verified Synthetic Web Environments for Agent LearningAugust 22, 2026 · arXiv
- How Mobile World Model Guides GUI Agents?May 11, 2026 · arXiv
- Learning with Challenges: Adaptive Difficulty-Aware Data Generation for Mobile GUI Agent TrainingJanuary 30, 2026 · arXiv
- Adapting Web Agents with Synthetic SupervisionNovember 8, 2025 · arXiv