Skip to main content
Glama

Batcave Detective Basic

Scenario Council

scenario_council

Runs exactly four fixed roles (Operator, Skeptic, Buyer, Adversary) for 1-3 bounded rounds over only the caller-supplied question, context, and evidence. Returns every role position, evidence lineage, structured peer challenges, belief changes, disagreements, falsifiers, deterministic process-integrity results, and one independent Detective Basic audit of the four final conclusions. Does not retrieve outside facts, execute actions, treat peer statements as evidence, or treat majority agreement as truth. Price: 0.25 USD via x402. Calling this MCP tool returns the canonical REST/x402 checkout contract; it does not execute the product or charge the caller yet.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
job_idNo
roundsNo
contextNo
evidenceNo
questionYes

Schema Changelog

Changes observed during successful MCP inspections.

  1. First observed

TDQS

A4.5/5.0
Behavior5/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the full behavioral burden and does so thoroughly. It discloses that calling returns a REST/x402 checkout contract, does not execute the product or charge the caller, does not retrieve outside facts, does not execute actions, and does not treat majority agreement as truth. It also surfaces deterministic process-integrity results and a price.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is dense but every sentence earns its place: action, return payload, exclusions, pricing, and checkout side-effect. The most important scoping information is front-loaded in the first sentence, and the rest is tightly organized.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a complex tool with no annotations and no output schema, the description covers action, inputs, output contents, constraints, cost, and side-effect behavior. An agent can understand what will happen, what it will receive, and what it will not be charged before invoking the tool.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 0%, so the description must compensate. It adds meaning for question, context, evidence, and rounds (via '1-3 bounded rounds'), but job_id is never mentioned and no parameter formats or constraints are described. Partial compensation with a clear gap for one optional parameter.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

Description opens with a specific verb and resource: 'Runs exactly four fixed roles (Operator, Skeptic, Buyer, Adversary) for 1-3 bounded rounds...' and clearly scopes inputs to caller-supplied question, context, and evidence. It further differentiates itself from execution-style siblings by stating it does not retrieve outside facts or execute actions, and returns a Detective Basic audit rather than being the detective tool itself.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description establishes clear context: use this tool for bounded multi-role analysis over caller-supplied material only, and it explicitly says it does not fetch outside facts or execute actions. It stops short of naming sibling tools or giving explicit when-to-use/when-not-to-use routing, so it misses the top score.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

Resources