Skip to content
Search
paperAugust 2026Unreviewed

TempJail: Temporal Jailbreak Attack against Large Vision-Language Models via Subtitle Scheduling

Ling Zhou, Yihao Huang, Jinglin Sun, Zhiwen Tian, Yi Zeng, Qi-He Liu, Shi-Jie Zhou

Abstract

Large vision-language models (LVLMs) have achieved remarkable progress in video understanding and reasoning. Despite extensive studies on text- and image-based jailbreaks, video jailbreaks against LVLMs remain largely unexplored. Existing video jailbreak methods mainly manipulate textual content embedded in videos, while overlooking how such information is organized over time. Our analysis reveals that jailbreak effectiveness depends not only on the semantics of textual information but also on i

Categories

Framework mappings

MITRE ATLAS
  • AML.T0054LLM Jailbreak

Suggested from the entry's categories.

Cite

@misc{zhou2026tempjail,
  title = {{TempJail: Temporal Jailbreak Attack against Large Vision-Language Models via Subtitle Scheduling}},
  author = {Ling Zhou and Yihao Huang and Jinglin Sun and Zhiwen Tian and Yi Zeng and Qi-He Liu and Shi-Jie Zhou},
  year = {2026},
  month = aug,
  eprint = {2608.19737},
  archivePrefix = {arXiv},
  url = {https://www.semanticscholar.org/paper/9256fdf3d5d0f55a454ebf4d96a30155cdf62b38}
}