Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the output schema exists, the description need not explain return values in depth, but it still covers key output elements: severity, why_flagged, review_questions, and flag types. It also places the tool in the Mode 4.3 workflow, so an agent has enough context to call and interpret it properly.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.