Newest first ยท 3 reviewed on this page
Search insteadreport/2025arXiv preprintReviewed Blake Bullwinkel, Amanda Minnich, Shiven Chawla +23
Shares lessons from Microsoft's AI red team operations including methodology, tooling, common failure modes, and best practices.
paper/2024arXiv preprintReviewed Sayash Kapoor, Rishi Bommasani, Kevin Klyman +2
Analyzes the societal impacts of open-weight foundation models, including security implications of open vs closed model access.
paper/2024Proc. SPIE 13054, Assurance and Security for AI-enabled SystemsReviewed Chris M. Ward, Josh Harguess, Julia Tao +3
Adapts David Bianco's Pyramid of Pain framework to AI security, categorizing AI threats by how difficult they are for adversaries to change.