prompt_injection_jailbreak_classifier
Classify user inputs for indirect prompt injections, delimiter hijacking, and system role impersonation to stop malicious instructions before they reach your LLM.
Instructions
High-speed deterministic classifier detecting indirect prompt injections, delimiter hijacking, and system role impersonation attacks in user inputs. (0.045 USDC on Base L2)
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| payload | Yes | Input parameters or JSON string payload for the tool execution | |
| paymentSignature | No | Base L2 USDC micropayment signature or transaction hash for x402 settlement |