Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
With no annotations, the description carries the full burden. It discloses a key behavioral trait: only hypotheses that passed the delivery evidence threshold are returned, and unverified ones are filtered out. This adds meaningful context beyond the schema, though it does not cover side effects, errors, or permissions.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.