scan_text
Scan text to detect prompt injection, data exfiltration, credential leaks, and other AI agent security threats, returning allow/block/quarantine decisions with severity and findings.
Instructions
Scan text for prompt injection attacks, data exfiltration attempts, credential leaks, and other AI agent security threats. Returns a decision (allow/block/quarantine), severity, and detailed findings.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| text | Yes | The text content to scan for security threats. | |
| channel | No | The input channel type. Affects which patterns are checked. Unknown channels are rejected (fail closed). | message |