Skip to main content
Glama

Ask Pipeworx — Grounded

ask_pipeworx_grounded
Read-onlyIdempotent

Hallucination-resistant answer mode for high-stakes reads. Same routing as ask_pipeworx — picks the right tool from 5,804 across 1518 sources, fills arguments, fetches the data — then EXTRACTS the answer using ONLY what the tool result contains. Returns {answer, evidence (verbatim quote), confidence, source, fetched_at, refusal_reason:null} on success, OR an explicit refusal {answer:null, refusal_reason:"not_in_source"|"no_tool_match"|"tool_error"|"data_truncated"|"llm_error"} when the data doesn't directly answer. Use whenever an answer will be quoted, cited, or acted on, and the agent must not invent facts (financial verdicts, legal claims, medical lookups, public statements). Costs one extra LLM call vs ask_pipeworx — prefer ask_pipeworx for casual lookups.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
qNoAlias for question.
textNoAlias for question.
inputNoAlias for question.
queryNoAlias for question.
promptNoAlias for question.
questionYesYour question in natural language. Accepts query, q, prompt, text, input as aliases.

Schema Changelog

Changes observed during successful MCP inspections. Dates show when Glama detected each change.

  1. First observed

TDQS

A4.7/5.0
Behavior5/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

The description goes well beyond the annotations by disclosing the refusal behavior, the exact refusal_reason enum values, and the structured success response. It also states that answers are extracted using ONLY the tool result, and that grounded mode costs an extra LLM call. These are important behavioral traits not captured by readOnlyHint, openWorldHint, idempotentHint, or destructiveHint.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is dense but well-organized, front-loading the core purpose, then explaining routing, return behavior, and usage guidance in a logical flow. Every sentence contributes essential information, and the cost comparison with ask_pipeworx earns its place. It is long only because it packs genuinely useful detail.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Despite lacking an output schema, the description fully specifies the success return shape, refusal shape, refusal reason categories, and high-stakes use cases. It also explains the relationship to ask_pipeworx and when to prefer the cheaper sibling. For a single-parameter tool with rich behavioral nuance, this is complete.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

The input schema already covers all six parameters, with 100% description coverage, documenting 'question' and its aliases. The description adds context about the question being routed and grounded, but does not add new parameter-level semantics. A baseline score of 3 is appropriate since the schema carries the parameter documentation burden.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly identifies the tool as a hallucination-resistant, grounded answer mode for high-stakes reads, using a specific verb ('extracts the answer') and a concrete resource. It explicitly distinguishes itself from ask_pipeworx by emphasizing that answers are drawn only from tool result content. This lets an agent immediately understand what the tool does and how it differs from its sibling.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines5/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description gives explicit when-to-use guidance: 'Use whenever an answer will be quoted, cited, or acted on' and lists high-stakes domains like financial verdicts, legal claims, medical lookups, and public statements. It also names ask_pipeworx as the alternative for casual lookups and explains the cost tradeoff. This fully routes an agent to the correct tool.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

TDQS

A3.6/5.0
Disambiguation1/5

The tool set is dominated by tools unrelated to USGS earthquakes (e.g., Polymarket betting, company profiles, memory operations). An agent would find it nearly impossible to distinguish the few earthquake-specific tools from the multitude of unrelated ones, leading to severe misselection.

Naming Consistency3/5

Most tool names follow a verb_noun pattern with underscores (e.g., search_earthquakes, count_earthquakes), which is consistent. However, the variety of verbs and domains creates a sense of incoherence, and some tool names are overly generic (e.g., process, run) in the broader context, though those are not present here. The naming pattern is acceptable but the inconsistency in domain scope reduces clarity.

Tool Count1/5

With 29 tools but only 3 directly related to earthquakes, the tool count is grossly inappropriate. The server's name suggests a focused purpose, but the vast majority of tools belong to other domains (e.g., Pipeworx queries, Polymarket betting, company data). This extreme mismatch makes the tool set bloated and misleading.

Completeness2/5

For earthquake data, the server provides only search, count, and get by ID. Missing are common operations like listing recent quakes, subscribing to alerts, or updating/correcting data. The coverage is minimal and insufficient for a comprehensive earthquake tool server, leaving significant gaps that agents could not work around.