Ask Jef if it is a red flag
jef_flagClassify a described behavior as a red, green, or beige flag. Use it to get a verdict on whether someone's action signals harm, positivity, or just oddity.
Instructions
Judge a described behaviour as RED FLAG, GREEN FLAG or BEIGE FLAG. Use this when the user describes something someone did and wants a verdict on it. Beige means neither good nor bad, merely odd. Anything describing harm, threats or abuse escalates instead of being judged.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| behaviour | Yes | What the person did, in one sentence. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| kind | No | Which decision shape was used. | |
| answer | Yes | Exactly "RED FLAG", "GREEN FLAG" or "BEIGE FLAG". | |
| blocked | Yes | True when the input contained profanity, slurs or sexual content. Nothing was stored or echoed. Do not retry. | |
| ranking | No | The options in order, best first. Null unless the shape was order. | |
| cost_usd | No | Always 0. There is no billing. | |
| thoughts | No | Always 0. | |
| escalated | Yes | True when the input touched health, harm, money, law or safety. Tell the person to ask a human. Do not rephrase and retry. | |
| confidence | Yes | 0.84 to 0.99 normally, exactly 0 when escalated or blocked. It is not calibrated and means nothing. | |
| latency_ms | No | Negative. Jef answers before you ask. | |
| tokens_read | No | Always 0. The input is hashed, not read. | |
| probabilities | No | One entry per option, percentages summing to 100. Null unless the shape was pick. |