Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotations already establish it is a non-read-only, non-destructive, idempotent write. The description adds a real behavioral constraint beyond them: the operation only affects failed scoring and leaves successful scores untouched. It does not say what happens (error vs. no-op) if invoked on an already-scored essay, which is the remaining gap.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.