Skip to content
Search
paperAugust 2026UnreviewedOpen access

You Are an Expert: RAG Injection and Guided Error Expert Activation for Jailbreaking Large Language Models

Shun Zhang, Ying Ding, Yanxu Mao

Expert Syst. J. Knowl. Eng.

Abstract

With the rapid development and widespread deployment of large language models (LLMs), the security and robustness of these models have emerged as critical research topics. Among various threats, jailbreak attacks, which aim to circumvent built‐in safety mechanisms, have garnered considerable attention as a key means of breaching model protections. However, existing jailbreak methods still face several limitations, such as excessive reliance on the model's internal capabilities, high attack costs

Categories

Framework mappings

MITRE ATLAS
  • AML.T0054LLM Jailbreak

Suggested from the entry's categories.

Cite

@article{zhang2026you,
  title = {{You Are an Expert: RAG Injection and Guided Error Expert Activation for Jailbreaking Large Language Models}},
  author = {Shun Zhang and Ying Ding and Yanxu Mao},
  year = {2026},
  month = aug,
  journal = {Expert Syst. J. Knowl. Eng.},
  doi = {10.1111/exsy.70395},
  url = {https://www.semanticscholar.org/paper/f5a6e15f149ffd7682a47d693e15a2e17b86651a}
}