← Back to search
paper llmsec-2026-00081
Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models
Zeyuan Chen, Yihan Ma, Xinyue Shen, Michael Backes, Yang Zhang
2026-05
Abstract
Large language models (LLMs) show strong performance across many applications, but their ability to memorize and potentially reveal training data raises serious privacy concerns. We introduce the PopQuiz Attack, a black-box membership inference attack that tests whether a model can recall specific training examples. The core idea is to turn target data into quiz-style multiple-choice questions and infer membership from the model's answers. Across six widely used LLMs (GPT-3.5, GPT-4o, LLaMA2-7b,
Categories
Cite This Resource
@article{llmsec202600081,
title = {Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models},
author = {Zeyuan Chen and Yihan Ma and Xinyue Shen and Michael Backes and Yang Zhang},
year = {2026},
url = {https://arxiv.org/abs/2605.06423},
} Metadata
- Added
- 2026-05-17
- Added by
- automation
- Source
- arxiv
- arxiv_id
- 2605.06423