← Back to search
paper llmsec-2026-00202

BadAgent Extension

Pedro Yanes Garrido, Diego Fernandez Arias

2026-06

Abstract

This paper presents an empirical study on backdoor attacks in large language model agents. We extend a recent attack framework by adding two lightweight benchmarks that measure cross-domain robustness and trigger visibility without changing the model architecture. Our approach fine-tunes AgentLM-based agents with parameter-efficient methods on operating system and web browsing tasks using multiple poisoning ratios and both visible and invisible triggers. We then evaluate the agents with

Cite This Resource

@article{llmsec202600202,
  title = {BadAgent Extension},
  author = {Pedro Yanes Garrido and Diego Fernandez Arias},
  year = {2026},
  doi = {10.4018/979-8-3373-8252-4.ch010},
  url = {https://doi.org/10.4018/979-8-3373-8252-4.ch010},
}

Metadata

Added
2026-05-17
Added by
automation
Source
crossref
doi
10.4018/979-8-3373-8252-4.ch010