Skip to content
Search
tool2023ReviewedOpen access

Rebuff: Self-Hardening Prompt Injection Detector

Protect AI

GitHub

Abstract

Open-source tool designed to detect and prevent prompt injection attacks using multiple detection methods including heuristics, LLM-based analysis, and canary tokens.

Categories

#tool#detection#canary-tokens#open-source

Framework mappings

Cite

@misc{protect2023rebuff,
  title = {{Rebuff: Self-Hardening Prompt Injection Detector}},
  author = {{Protect AI}},
  year = {2023},
  url = {https://github.com/protectai/rebuff}
}