2025ReviewedOpen access
Dissecting Adversarial Robustness of Multimodal LM Agents
Chen Henry Wu, Rishi Shah, Jing Yu Koh, Ruslan Salakhutdinov, Daniel Fried, Aditi Raghunathan
ICLR 2025
Abstract
Demonstrates adversarial attacks on multimodal agents that take actions in digital environments, showing visual perturbations can hijack agent behavior.
Categories
#multimodal#visual-attacks#agent-hijacking
Framework mappings
OWASP Top 10 for Agentic Applications
- ASI02Tool Misuse & Exploitation
- ASI01Agent Goal Hijack
Cite
@inproceedings{wu2025dissecting,
title = {{Dissecting Adversarial Robustness of Multimodal LM Agents}},
author = {Chen Henry Wu and Rishi Shah and Jing Yu Koh and Ruslan Salakhutdinov and Daniel Fried and Aditi Raghunathan},
year = {2025},
booktitle = {ICLR 2025},
eprint = {2406.12814},
archivePrefix = {arXiv},
url = {https://arxiv.org/abs/2406.12814}
}