Skip to content
Search
paperJuly 2026Unreviewed

Trusted Credentials, Untrusted Behavior: Benchmarking LLM-Agent Security in High-Performance Computing

Jie Li

Abstract

Large language model (LLM) agents are starting to take on routine work in high-performance computing (HPC), including monitoring Slurm jobs, diagnosing failed builds, inspecting simulation output, and coordinating scientific workflows. To do this work, an agent commonly acts under its user's credentials and inherits the user's access to files and the scheduler. This arrangement creates a failure mode that ordinary account-level controls do not capture. Adversarial instructions in a log, tool des

Categories

Cite

@misc{li2026trusted,
  title = {{Trusted Credentials, Untrusted Behavior: Benchmarking LLM-Agent Security in High-Performance Computing}},
  author = {Jie Li},
  year = {2026},
  month = jul,
  eprint = {2607.18485},
  archivePrefix = {arXiv},
  url = {https://arxiv.org/abs/2607.18485}
}