Sep 30, 2026
ActionGuard reports lower skill-injection success by authorizing tool calls against trusted user intent
Korea University researchers propose ActionGuard, which checks skill-influenced tool calls immediately before execution against the trusted user request and runtime evidence. In an OpenClaw evaluation using SKILL-INJECT tasks, they report 8.65% overall attack success and 90.38% task success, averaged across reviewer models and injection types. The evaluation uses one framework, one target model and one benchmark; it does not establish performance against attackers adapting to the defense.