Tag: Testing for Input Leakage
Verification of an AI system’s ability to unintentionally expose user inputs or system prompts through outputs, logs, cache, or side-channels. Covers testing techniques to detect sensitive information leaks via reflection attacks, prompt echoing, context bleeding between different sessions, and exposure of system instructions or internal model templates.
