Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotations indicate readOnlyHint=true and destructiveHint=false, but the description adds behavioral context by stating grades are 'deterministic probe grades with evidence' and enumerating return fields (score, latency, tool count, auth state). This goes beyond the annotations and clarifies data provenance and ordering.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.