Skip to main content
Glama

nsgoods-workbench-mcp

Host summary

host_summary
Read-onlyIdempotent

Summary for a host (no time series): first/last seen, current verdict mix, n_resources, total verdict changes, gone_since if absent from the latest full scan, and a small resources_sample. Sample rows with in_latest_scan=false were not in the latest full scan; read last_checked_at for when they were last probed. n_resources counts every resource ever seen for this host; current_verdict_mix and n_payable_last_full count only the latest full scan. last_checked_at is the last time any scan looked at the resource; verdict_since is when the current verdict was first observed in the unbroken run that leads to it. A scan that did not contain the resource does not break the run, so verdict_since can span weeks nobody measured. How current a row is comes from last_checked_at and in_latest_scan, not from verdict_since alone.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
hostYes

Schema Changelog

Changes observed during successful MCP inspections.

  1. First observed

TDQS

B3.3/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations already declare readOnlyHint, idempotentHint, and destructiveHint=false, so the safety profile is covered. On top of that the description discloses genuinely non-obvious semantics: n_resources counts resources ever seen while current_verdict_mix and n_payable_last_full count only the latest full scan, and verdict_since can span unmeasured weeks because a scan missing the resource does not break the run. That is real behavioral context an agent cannot derive from annotations.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

With no output schema, the field-by-field explanation is load-bearing rather than padding, and it is front-loaded with the purpose and field inventory before the caveats. A few clauses restate the same latest-full-scan distinction (n_resources vs current_verdict_mix/n_payable_last_full), which costs some tightness, but overall every sentence earns its place.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Because there is no output schema, the description carries the full burden of explaining returned fields and does so thoroughly, including the subtle veracity caveat about verdict_since. The one real hole is the undocumented 'host' argument, which leaves an agent guessing at input format. For an otherwise complex read-only summarizer, it is close to complete.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters2/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

There is one required parameter, 'host', and schema description coverage is 0%. The description never says what form the host argument takes (hostname, identifier, URL fragment) or how to obtain it, so the single parameter is undocumented in both the schema and the prose. With one parameter and no coverage, the description should have compensated and does not.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

The first clause states a specific verb and resource ('Summary for a host') and explicitly bounds scope with '(no time series)', which is a meaningful disambiguation from time-series siblings. It immediately enumerates the returned fields, so an agent knows exactly what it gets. It does not name a sibling tool to contrast against, but the sibling set (drift_status, payability_verdict, etc.) does not obviously overlap, so this is a minor gap.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines2/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Nothing states when to call this versus alternatives like payability_verdict or catalogue_stats, nor any prerequisites such as whether a scan must exist first. The guidance about interpreting last_checked_at vs verdict_since is about reading results, not about choosing the tool. Usage is only implied by the tool's name and scope.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.