tos_watch_recent_changes
Recent revised events across all SaaS ToS documents since the given ISO8601 timestamp. Each item includes firstSeenAt and ledgerVerified.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| limit | No | ||
| since | Yes | ||
| vendor | No |
Recent revised events across all SaaS ToS documents since the given ISO8601 timestamp. Each item includes firstSeenAt and ledgerVerified.
| Name | Required | Description | Default |
|---|---|---|---|
| limit | No | ||
| since | Yes | ||
| vendor | No |
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
The description discloses that results include 'firstSeenAt' and 'ledgerVerified', providing output context beyond the schema. However, it does not state that the tool is read-only, nor does it mention rate limits, authentication, or any side effects. With no annotations, the description carries the full burden but only partially addresses it.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
Two sentences, front-loaded with the core purpose, followed by a useful output detail. Every word adds value with no redundancy. Highly concise.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the tool has three parameters, no output schema, and no annotations, the description is incomplete. It omits 'limit' and 'vendor' explanations, does not describe the return format (list? pagination?), and lacks behavioral context. Significant gaps remain for an agent to use it effectively.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
The description explains the 'since' parameter (ISO8601 timestamp) but ignores 'limit' and 'vendor'. With 0% schema description coverage, the description was expected to compensate but only covers one of three parameters, leaving significant gaps for the agent.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
Clearly states it retrieves recent revised events of SaaS ToS documents since a timestamp. The verb 'watch recent changes' is specific, and the resource is identified. However, it does not explicitly differentiate from sibling tools like tos_watch_search or other domain watch tools, though the name and context make it distinct enough.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
No guidance on when to use this tool versus alternatives such as tos_watch_search or tos_watch_timeline. The description lacks any context about prerequisites, exclusions, or typical use cases, leaving the agent to infer usage.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
Add one secure layer between your agents and this server.
Most tools have clearly distinct purposes, especially the ledger-watch groups with domain prefixes. However, there are near-duplicate utilities such as content_authenticity_domain_reputation and domain_intel_reputation, and fx_tax_convert overlaps with price_oracle_fx_rate/convert. The sheer number of tools also increases the chance of selecting the wrong one, though descriptions are generally clear.
The naming is largely consistent with a <domain>_<action> or <domain>_<watch>_<action> pattern, and all names use snake_case. Minor deviations include standalone names like entity_search, kyb_report, verify_receipt, and the confusing singular/plural pair of sanction_watch_* and sanctions_screen_*. Overall, the pattern is predictable.
With 172 tools, this server is far beyond the typical well-scoped range. It bundles dozens of unrelated utility domains (weather, carbon, CVE, geo, etc.) alongside the Japan public-ledger watches, making it unwieldy for an agent to navigate. This extreme count is a major coherence problem.
For the core Japan public-ledger domain, coverage is excellent: each ledger has search, get, timeline, recent_changes, and verify_ledger, plus cross-ledger entity_search and temporal_query. The unrelated utility areas are also fairly complete for their own purposes, but the server's scope is so broad that some utilities are duplicated (e.g., multiple currency converters). Overall, no critical dead ends in the main domain.