Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
There are no annotations, so the description carries the behavioral burden. It adds meaningful context by specifying that LAB benchmarks, tests, shadow runs, and replay events are excluded, making clear the tool returns only production aggregate data. Given the output schema exists, the lack of return-format details is acceptable.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.