Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the tool's complexity (multiple invocation modes, sibling distinctions, output behavior), the description covers all critical aspects: common user phrasings, parameter constraints (max 12), output format, and a clear pointer for videos. With an output schema present, return values are handled elsewhere, so nothing essential is missing.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.