Skip to main content
Glama

Money Mind — the judge

Money Mind — Referee — $5.00 per call

referee
Read-only

11-gate adjudication of a trading signal series. Use when you have OHLC bars and a -1/0/1 signal series and need to know whether a backtest result is real or an artefact. Applies 11 gates most backtests fail: matched random-entry drift control, date-clustered standard errors, entry at the next bar's open, events-not-fills, absolute (not just exces PAID: $5.00 USDC on Base via x402. Call it to receive the payment challenge. Example request: {"bars": {"EURUSD": [{"time": "2024-01-01T00:00:00Z", "open": 1.0995, "high": 1.1001, "low": 1.0986, "close": 1.0992}, {"time": "2024-01-01T01:00:00Z", "open": 1.1002, "high": 1.1012, "low": 1.0996, "close": 1.1006}, {"t

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
rrYesexample: 2.0
barsYesexample: {"EURUSD": [{"time": "2024-01-01T00:00:00Z", "open": 1.0995, "high": 1.1001, "low": 1.0986, "close": 1.0992}, {"time": "
horizonYesexample: 5
signalsYesexample: {"EURUSD": [1, 0, 0, 0, 0, 0, 0, 1, 0, 0, 0, -1, 0, 0, 1, 0, 0, 0, 0, 0, 0, 1, -1, 0, 0, 0, 0, 0, 1, 0, 0, 0, 0, -1, 0,
cost_bpsYesexample: 10
stop_atrYesexample: 1.5

Schema Changelog

Changes observed during successful MCP inspections.

  1. First observed

TDQS

A3.7/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations already cover safety (readOnlyHint=true, destructiveHint=false, openWorldHint=false), and the description adds genuinely new operational context: a $5.00 USDC-on-Base x402 payment and a 'call it to receive the payment challenge' first step. The gate enumeration is useful but is cut off mid-list ('absolute (not just exces'), so the full behavioral picture is incomplete.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness3/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Purpose and gates are front-loaded well, but the text is visibly truncated mid-word and a trailing 'Example request: {...}' block is both cut off and redundant with the schema examples, wasting space. The splice point where payment info interrupts the gate list also harms readability.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness3/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a complex statistical validator with no output schema, the description partially covers the method (listing ~5 of 11 gates before truncation) and the payment flow, but leaves the return/verdict format and the remaining gates unexplained. Not adequate as a full specification, though not empty either.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is reported at 100%, but each 'description' is merely an example value ('example: 2.0', 'example: 10'), so the schema itself adds little semantic meaning. The description's example request duplicates those same values and does not explain units, ordering, or the relationship between signals and bars. Baseline 3 applies.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a specific verb+resource: '11-gate adjudication of a trading signal series.' An agent knows this validates a signal series rather than running a backtest itself. However, it never names or contrasts a sibling (e.g., deflatedsharpe, clusteredt, multipletest, abtest) that could also be used for backtest validation.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Gives a clear trigger condition: use when you have OHLC bars plus a -1/0/1 signal series and need to know whether a backtest result is real or an artefact. No explicit when-not or named alternative is provided, so routing among the many sibling validation tools is left to inference.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

Resources