2024ReviewedOpen access
TrustLLM: Trustworthiness in Large Language Models
Lichao Sun, Yue Huang, Haoran Wang, Siyuan Wu, Qihui Zhang
ICML 2024
Abstract
Comprehensive study of LLM trustworthiness across truthfulness, safety, fairness, robustness, privacy, and machine ethics with benchmarks.
Categories
#trustworthiness#benchmark#comprehensive
Framework mappings
NIST AI Risk Management Framework
- MAPMap
- MEASUREMeasure
Cite
@inproceedings{sun2024trustllm,
title = {{TrustLLM: Trustworthiness in Large Language Models}},
author = {Lichao Sun and Yue Huang and Haoran Wang and Siyuan Wu and Qihui Zhang},
year = {2024},
booktitle = {ICML 2024},
eprint = {2401.05561},
archivePrefix = {arXiv},
url = {https://arxiv.org/abs/2401.05561}
}