arXiv Artificial Intelligence

MMSkillRisk: Can Agents Stay Safe When Multimodal Skills Become Traps?

MMSkillRisk: Can Agents Stay Safe When Multimodal Skills Become Traps?

Quick summary

arXiv:2609.35912v1 Announce Type: cross Abstract: Agent skills are shareable packages of procedural instructions, tools, and examples. Multimodal skills additionally include visual references that agents retrieve and inspect during execution. Because these images guide actions, attackers can disguise malicious instructions as ordinary visual guidance within otherwise legitimate skills. Existing skill-security research primarily examines text-carried attacks or scanner detection, leaving the runtime effects of image-borne attacks insufficiently evaluated. We introduce MMSkillRisk, to our knowle

Key takeaways

  • arXiv:2609.35912v1 Announce Type: cross Abstract: Agent skills are shareable packages of procedural instructions, tools, and examples.
  • Multimodal skills additionally include visual references that agents retrieve and inspect during execution.
  • Because these images guide actions, attackers can disguise malicious instructions as ordinary visual guidance within otherwise legitimate skills.

Why it matters

“MMSkillRisk: Can Agents Stay Safe When Multimodal Skills Become Traps?” should be evaluated beyond branding and benchmark scores. Its practical importance will emerge in task accuracy, latency, unit cost, safety and integration with real workflows.

Kaynak sitede devamını oku: arXiv Artificial Intelligence ↗