← Back to search
paper llmsec-2026-00134

BadDLM: Backdooring Diffusion Language Models with Diverse Targets

Shengfang Zhai, Xiaoyang Ji, Yuling Shi, Haoran Gao, Fanyu Meng, Yan Zeng, Yuejian Fang, Yinpeng Dong, Jiaheng Zhang

2026-05

Abstract

Diffusion language models (DLMs) have recently emerged as an alternative modeling paradigm to autoregressive (AR) language models, enabling parallel generation and bidirectional context modeling. Yet their security implications, particularly their vulnerability to backdoor attacks, remain underexplored. We propose BadDLM, a unified framework for studying backdoor attacks against DLMs with diverse targets. We introduce a trigger-aware training objective that emphasizes target-relevant positions i

Categories

Cite This Resource

@article{llmsec202600134,
  title = {BadDLM: Backdooring Diffusion Language Models with Diverse Targets},
  author = {Shengfang Zhai and Xiaoyang Ji and Yuling Shi and Haoran Gao and Fanyu Meng and Yan Zeng and Yuejian Fang and Yinpeng Dong and Jiaheng Zhang},
  year = {2026},
  url = {https://arxiv.org/abs/2605.09397},
}

Metadata

Added
2026-05-17
Added by
automation
Source
arxiv
arxiv_id
2605.09397