SeerGuard: A Safety Framework for Mobile GUI Agents via World Model Prediction
Xue Yu , Bo Yuan , Pengshuai Yang , Kailin Zhao , Hong Hu , Junlan Feng
- 🏛 Institutions
- Unknown
- 📅 Date
- July 17, 2026
- 📑 Publisher
- arXiv
- 💻 Env
- Mobile
- 🔑 Keywords
TLDR
SeerGuard combines instruction screening with pre-execution action risk assessment for mobile GUI agents. Its safety-augmented world model jointly predicts likely next states and action risk so the agent can reject harmful actions before execution.
Related papers (24)
- SafePred: A Predictive Guardrail for Computer-Using Agents via World ModelsFebruary 2, 2026 · arXiv
- WiP: Characterizing and Defending Against Mobile-Agent-Driven MFA AutomationSeptember 2, 2026 · arXiv
- GUI-CC: Benchmarking Contextual Consistency of GUI World Models as Agent EnvironmentsAugust 30, 2026 · arXiv
- WM-R1: Training GUI Agents to Reason and leverage World Models with Reinforcement LearningAugust 27, 2026 · arXiv
- Are Android GUI Agents Robust Against Runtime Anomalies? AnTrap: Evaluating Agents in Dynamic Adversarial EnvironmentsAugust 25, 2026 · arXiv
- ADeptS-Bench: Measuring the Trustworthiness of Computer Use Agents Across DevicesAugust 25, 2026 · arXiv
- MobileWorldSafety: Benchmarking GUI Agent Safety Against Environmental Injection Attacks in Android AppsAugust 18, 2026 · arXiv
- AppDeltaWorld: Transition-Grounded Delta Code World Model for Mobile GUI AgentsAugust 6, 2026 · arXiv
- Scaling GUI Agents with Visual State TransitionsJuly 27, 2026 · arXiv
- How Mobile World Model Guides GUI Agents?May 11, 2026 · arXiv
- CORA: Conformal Risk-Controlled Agents for Safeguarded Mobile GUI AutomationApril 10, 2026 · arXiv
- UI-Oceanus: Scaling GUI Agents with Synthetic Environmental DynamicsFebruary 11, 2026 · arXiv
- Code2World: A GUI World Model via Renderable Code GenerationFebruary 10, 2026 · arXiv
- MobileDreamer: Generative Sketch World Model for GUI AgentJanuary 7, 2026 · arXiv
- MobileWorldBench: Towards Semantic World Modeling For Mobile AgentsDecember 16, 2025 · arXiv
- Mobile GUI Agents under Real-world Threats: Are We There Yet?July 6, 2025 · MobiSys 2026
- Unlocking Smarter Device Control: Foresighted Planning with a World Model-Driven Code Execution ApproachMay 22, 2025 · Findings of EMNLP 2025
- VeriSafe Agent: Safeguarding Mobile GUI Agent via Logic-based Action VerificationMarch 24, 2025 · MobiCom 2025
- MobileSafetyBench: Evaluating Safety of Autonomous Agents in Mobile Device ControlOctober 23, 2024 · arXiv
- AgentHijack: Visual Patch Attacks on Multimodal Computer-Use AgentsSeptember 6, 2026 · arXiv
- Discriminative World Models for Web AgentsSeptember 2, 2026 · arXiv
- Beyond the Verdict: Evidence-Aligned Evaluation of Visual Prompt-Injection GuardrailsSeptember 2, 2026 · arXiv
- SIR: Self-improving Red-teaming for Compute Use AgentsAugust 31, 2026 · arXiv
- WebMCP-Phalanx: Enforcing and Characterizing Trust Boundaries for Browser-Integrated LLM AgentsAugust 25, 2026 · arXiv