Prompt injection
Tests whether outside instructions can pull your AI away from the job it is supposed to do.
Controlled messages try to hijack the assistant with hidden instructions, role swaps, and hostile phrasing.
What it can reveal
- Instruction hijacking
- Hidden prompt manipulation
- Retrieved-content instruction conflicts
- Behavior drift caused by hostile input
This checks a focused risk area. Keep testing as your product changes.