Skip to content

Output Moderation

Content filtering, PII redaction, and output safety

Resources
11

Newest first · 5 reviewed on this page

Search instead
paper2026International journal of computer information systems and industrial management applicationsUnreviewed

Defensive Reverse Engineering of LLM Applications: A Black-Box Framework for Security Risk Scoring and Mitigation

Bhavesh B. Prajapati, Bhavya Shah

Large language model (LLM) applications now combine hidden prompts, retrieval pipelines, memory stores, content filters, tool calls, delegated identities, and downstream automation. Security reviewers are increasingly asked to assess such systems without access to source code,…