Prompt Shield (Jailbreak / Injection Detection)
detect_prompt_injectionClassify a prompt before it reaches your LLM. Brainiall Prompt Shield engine.
Returns category (jailbreak | prompt_injection | data_exfiltration | impersonation | none), severity, reason, confidence.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| prompt | Yes | The prompt text to classify (NOT executed) |