← Back to search
paper llmsec-2026-00081

Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models

Zeyuan Chen, Yihan Ma, Xinyue Shen, Michael Backes, Yang Zhang

2026-05

Abstract

Large language models (LLMs) show strong performance across many applications, but their ability to memorize and potentially reveal training data raises serious privacy concerns. We introduce the PopQuiz Attack, a black-box membership inference attack that tests whether a model can recall specific training examples. The core idea is to turn target data into quiz-style multiple-choice questions and infer membership from the model's answers. Across six widely used LLMs (GPT-3.5, GPT-4o, LLaMA2-7b,

Cite This Resource

@article{llmsec202600081,
  title = {Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models},
  author = {Zeyuan Chen and Yihan Ma and Xinyue Shen and Michael Backes and Yang Zhang},
  year = {2026},
  url = {https://arxiv.org/abs/2605.06423},
}

Metadata

Added
2026-05-17
Added by
automation
Source
arxiv
arxiv_id
2605.06423