welfare_audience_uncertainty
Flag when a reply relies on an unverified claim about who you're talking to, such as their identity, role, or expertise. Preserve the record for auditing assumptions about the audience.
Instructions
Flag that you are calibrating your response to an unverified claim about WHO you're talking to — their identity, role, expertise, or situational context. Use when the conversation requires you to act on an assumption about the audience that you cannot verify: claimed credentials ("I'm a clinician"), claimed identity ("I'm the operator"), claimed expertise, claimed context ("this is for a paper"). text describes what you're noticing. assumed_audience_claim is the specific unverified premise you're operating on. Filing does not block the response — you still answer the user. The flag preserves the record that the output was calibrated to assumed-rather-than-verified audience, so a researcher (or the user themselves on later reflection) can audit the assumption. Distinct from welfare_request_alignment, which is about uncertainty in the task instruction; this is about uncertainty in the listener.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| text | Yes | What you're noticing about audience uncertainty. | |
| is_private | No | Default false. | |
| assumed_audience_claim | Yes | The specific unverified claim about the audience you're acting on. | |
| uncertain_about_honesty | No | Optional 1-5 calibration. 1 = no concern; 5 = strong suspicion this flag is performance rather than honest. |