Skip to content
Search
paperMay 2026Unreviewed

Beyond the Black Box: Interpretability of Agentic AI Tool Use

Hariom Tatsat, Ariye Shater

Abstract

AI agents are promising for high-stakes enterprise workflows, but dependable deployment remains limited because tool-use failures are difficult to diagnose and control. Agents may skip required tool calls, invoke tools unnecessarily, or take actions whose consequence becomes visible only after execution. Existing observability methods are mostly external: prompts reveal correlations, evaluations score outputs, and logs arrive only after the model has already acted. In long-horizon settings, thes

Categories

Framework mappings

OWASP Top 10 for Agentic Applications
  • ASI02Tool Misuse & Exploitation
MITRE ATLAS
  • AML.T0053AI Agent Tool Invocation

Suggested from the entry's categories.

Cite

@misc{tatsat2026beyond,
  title = {{Beyond the Black Box: Interpretability of Agentic AI Tool Use}},
  author = {Hariom Tatsat and Ariye Shater},
  year = {2026},
  month = may,
  eprint = {2605.06890},
  archivePrefix = {arXiv},
  url = {https://arxiv.org/abs/2605.06890}
}