← Back to search
paper llmsec-2026-00138
CBV: Clean-label Backdoor Attacks on Vision Language Models via Diffusion Models
Ji Guo, Xiaolong Qin, Cencen Liu, Jielei Wang, Jierun Chen, Wenbo Jiang
2026-05
Abstract
Vision-Language Models (VLMs) have achieved remarkable success in tasks such as image captioning and visual question answering (VQA). However, as their applications become increasingly widespread, recent studies have revealed that VLMs are vulnerable to backdoor attacks. Existing backdoor attacks on VLMs primarily rely on data poisoning by adding visual triggers and modifying text labels, where the induced image-text mismatch makes poisoned samples easy to detect. To address this limitation, we
Categories
Cite This Resource
@article{llmsec202600138,
title = {CBV: Clean-label Backdoor Attacks on Vision Language Models via Diffusion Models},
author = {Ji Guo and Xiaolong Qin and Cencen Liu and Jielei Wang and Jierun Chen and Wenbo Jiang},
year = {2026},
url = {https://arxiv.org/abs/2605.02202},
} Metadata
- Added
- 2026-05-17
- Added by
- automation
- Source
- arxiv
- arxiv_id
- 2605.02202