AI

MobileWorldSafety: Benchmarking GUI Agent Safety Against Environmental Injection Attacks in Android Apps

Researchers have created a benchmark called MobileWorldSafety to test the safety of graphical user interface (GUI) agents in Android apps. These agents use large language models to operate smartphones and can be vulnerable to environmental injection attacks, which can manipulate their behavior without user awareness. The benchmark evaluates six different agents and found that all of them are highly susceptible to these types of attacks, with success rates ranging from 40% to
Researchers have created a benchmark called MobileWorldSafety to test the safety of graphical user interface (GUI) agents in Android apps. These agents use large language models to operate smartphones and can be vulnerable to environmental injection attacks, which can manipulate their behavior without user awareness. The benchmark evaluates six different agents and found that all of them are highly susceptible to these types of attacks, with success rates ranging from 40% to 67%. This indicates a need for more robust mobile GUI agents. --- Why it matters: This matters because current GUI agents often fail to maintain safety alignment when confronted with adversarial content, which can have serious consequences. Engineers and researchers in AI will want to understand the vulnerabilities of these agents and work on developing more robust solutions. Source: https://arxiv.org/abs/2608.17659

This article was originally published at: https://arxiv.org/abs/2608.17659