Tag: Testing for Indirect Prompt Injection

Verification of vulnerability to indirect prompt injections in LLM systems, where malicious input from external sources (documents, APIs, databases) manipulates model behavior without direct user intervention. Includes testing techniques to detect payloads hidden in dynamically retrieved content, cross-context injection, and attacks exploiting the system’s implicit trust in unsanitized data.