Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotations indicate this is a non-read, idempotent, non-destructive operation. The description adds that it keeps the original test's URL, goal, prompt, and inbox config, providing useful context beyond annotations. Minor gap: it doesn't specify whether a new job ID is created each time.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.