Skip to main content
Glama

Taifoon coordination layer

taifoon_grid_bench

Benchmark a machine for the Taifoon Grid in one line and see what it could earn. Returns the command to run ON the box (curl -fsSL https://www.taifoon.io/grid/bench.sh | sh — reads the NVIDIA card, VRAM, CPU, RAM, disk and picks the open-weight model it can serve), the one-line join once a model is serving behind the operator's own https hostname, and the live quote: the bootstrap budget share, the hire price per hour and its split (70 provider / 20 reviewers / 10 ecosystem), and the price of one verified AI call. Give vram_gb and gpus to get the model tier without running anything. Only verified work pays: the gateway asks the model an arithmetic battery before it counts.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
gpusNonumber of cards (default 1)
ownerNooptional 0x wallet — filled into the join line
vram_gbNoVRAM of one card in GB (0 or omitted = no GPU)

Schema Changelog

Changes observed during successful MCP inspections.

  1. First observed

TDQS

A3.7/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the full burden, and it discloses meaningful behavior: the exact shell command that runs on the box, what it inspects (NVIDIA card, VRAM, CPU, RAM, disk), the economic split (70/20/10), and the verification gate ("the gateway asks the model an arithmetic battery before it counts"). It does not state that the tool itself is a read-only/informational operation or note any auth or rate-limit requirements, so it stops short of a 5.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The lead clause is front-loaded and states the payoff ("see what it could earn"). The single dense paragraph is long but nearly every clause adds distinct information (command, join line, split, verification). Slightly heavy for a definition of this kind.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

There is no output schema, so the description must describe the returns, and it does so concretely: the curl one-liner, the join line, and the quote fields including the 70/20/10 split and per-call price. Combined with the verification caveat, an agent has enough to call it correctly, though the lack of any read-only/auth framing leaves a minor gap.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 100%, so the schema already documents gpus, vram_gb and owner. The description adds the behavioral meaning of vram_gb/gpus (they yield the model tier without running anything) but says nothing about the owner wallet beyond the schema's own note, so baseline 3 applies.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a specific verb and resource ("Benchmark a machine for the Taifoon Grid") and enumerates exactly what it returns (bench command, join line, live quote with splits and per-call price). An agent can tell it produces an economics/preview package rather than performing a join or querying prices. It does not explicitly name the sibling tools it overlaps with (taifoon_grid_join, taifoon_grid_prices), so it falls short of a 5.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines3/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

It gives one conditional usage hint ("Give vram_gb and gpus to get the model tier without running anything"), which is genuinely useful. However it never says when to prefer this over the sibling tools that cover overlapping ground (grid_join, grid_prices, node_command), leaving the routing decision to inference.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

Resources