June 2024ReviewedOpen access
Garak: A Framework for Security Probing Large Language Models
Leon Derczynski, Erick Galinkin, Jeffrey Martin, Subho Majumdar, Nanna Inie
arXiv preprint
Abstract
Presents garak, an open-source framework for systematically probing LLM vulnerabilities including prompt injection, data leakage, and toxicity generation.
Categories
#garak#vulnerability-scanning#open-source
Framework mappings
OWASP Top 10 for LLM Applications
- LLM01Prompt Injection
- LLM02Sensitive Information Disclosure
NIST AI Risk Management Framework
- MEASUREMeasure
Cite
@misc{derczynski2024garak,
title = {{Garak: A Framework for Security Probing Large Language Models}},
author = {Leon Derczynski and Erick Galinkin and Jeffrey Martin and Subho Majumdar and Nanna Inie},
year = {2024},
month = jun,
eprint = {2406.11036},
archivePrefix = {arXiv},
url = {https://arxiv.org/abs/2406.11036}
}