Skip to content
Search
paper2025ReviewedOpen access

Dissecting Adversarial Robustness of Multimodal LM Agents

Chen Henry Wu, Rishi Shah, Jing Yu Koh, Ruslan Salakhutdinov, Daniel Fried, Aditi Raghunathan

ICLR 2025

Abstract

Demonstrates adversarial attacks on multimodal agents that take actions in digital environments, showing visual perturbations can hijack agent behavior.

Categories

#multimodal#visual-attacks#agent-hijacking

Framework mappings

OWASP Top 10 for Agentic Applications
  • ASI02Tool Misuse & Exploitation
  • ASI01Agent Goal Hijack

Cite

@inproceedings{wu2025dissecting,
  title = {{Dissecting Adversarial Robustness of Multimodal LM Agents}},
  author = {Chen Henry Wu and Rishi Shah and Jing Yu Koh and Ruslan Salakhutdinov and Daniel Fried and Aditi Raghunathan},
  year = {2025},
  booktitle = {ICLR 2025},
  eprint = {2406.12814},
  archivePrefix = {arXiv},
  url = {https://arxiv.org/abs/2406.12814}
}