Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
With no annotations, the description carries the full burden and does add real behavioral context: every result carries a credibility score, Chinese queries are auto-translated, and before_chatgpt acts as a 2022-11-30 cutoff filter. It still omits return format, pagination, and any rate/auth expectations, keeping it short of a 5.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.