Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotations already declare readOnlyHint=true and openWorldHint=false, so the safety profile is covered. The description adds which metric families are compared (accuracy, cost, latency, tokens; overall and per-size), which is useful, but since an output schema exists this is largely a preview of return contents rather than disclosure of behavior an agent couldn't infer.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.