Tag: Testing for Prompt Injection
Prompt injection is an attack technique against Large Language Model (LLM) based systems that allows manipulating model behavior by inserting malicious instructions into the user prompt. Testing verifies if an AI application is vulnerable to crafted inputs that overwrite, bypass, or alter system instructions, causing unauthorized outputs, data leakage, or unexpected action execution. Includes direct injection techniques, indirect injection via external content, and jailbreak to circumvent security policies.
