Skip to main content
Glama

Ask Pipeworx — Grounded

ask_pipeworx_grounded
Read-onlyIdempotent

Hallucination-resistant answer mode for high-stakes reads. Same routing as ask_pipeworx — picks the right tool from 5,718 across 1496 sources, fills arguments, fetches the data — then EXTRACTS the answer using ONLY what the tool result contains. Returns {answer, evidence (verbatim quote), confidence, source, fetched_at, refusal_reason:null} on success, OR an explicit refusal {answer:null, refusal_reason:"not_in_source"|"no_tool_match"|"tool_error"|"data_truncated"|"llm_error"} when the data doesn't directly answer. Use whenever an answer will be quoted, cited, or acted on, and the agent must not invent facts (financial verdicts, legal claims, medical lookups, public statements). Costs one extra LLM call vs ask_pipeworx — prefer ask_pipeworx for casual lookups.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
qNoAlias for question.
textNoAlias for question.
inputNoAlias for question.
queryNoAlias for question.
promptNoAlias for question.
questionYesYour question in natural language. Accepts query, q, prompt, text, input as aliases.

TDQS

A4.7/5.0
Behavior5/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

The annotations already declare readOnlyHint, openWorldHint, idempotentHint, and destructiveHint=false, but the description adds meaningful behavioral context beyond these: it reveals the extraction-only constraint, the exact refusal reasons, the extra LLM call cost, and the structured return shape. This goes well beyond what annotations alone convey.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is relatively long but densely informative: it front-loads the core purpose, then explains behavior, return values, refusal cases, use cases, and cost trade-off. Every sentence contributes actionable information for selecting and invoking the tool correctly. No filler is present.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Despite having no output schema, the description fully documents the return payload and explicit refusal reasons, which is essential for a tool that must not hallucinate. It covers success and failure modes, use cases, cost implications, and routing behavior. With rich annotations and 100% parameter coverage, nothing critical is missing.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

The input schema already has 100% description coverage, explaining that all six parameters are aliases for the natural-language question. The tool description does not add any further parameter-level semantics beyond mentioning that the tool 'fills arguments' internally. Baseline 3 is appropriate because the schema carries the load.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly states this is a hallucination-resistant, grounded answer mode for high-stakes reads, and explains it routes to tools like ask_pipeworx but extracts answers strictly from tool results. It explicitly distinguishes itself from ask_pipeworx by emphasizing evidence and refusal behavior. A specific verb, resource, and unique behavior are all present.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines5/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description explicitly says to use this tool whenever an answer will be quoted, cited, or acted on and the agent must not invent facts, while also instructing to prefer ask_pipeworx for casual lookups due to the extra LLM call cost. This gives clear when-to-use and when-not-to-use guidance relative to the sibling tool.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

TDQS

A4/5.0
Disambiguation3/5

Most tools have clearly distinct roles, and the extensive descriptions help differentiate intent, but the ask_pipeworx family—especially ask_pipeworx_beta, which is currently identical to ask_pipeworx—creates real ambiguity. Overlapping entry points like ask_pipeworx, ask_pipeworx_grounded, deep_research, and validate_claim could also cause misselection without careful reading.

Naming Consistency4/5

All 34 tools use consistent snake_case, and clear verb-led or noun-prefixed patterns emerge across families like ask_pipeworx*, polymarket_*, and remember/recall/forget. Minor deviations such as ai_visibility_check and recent_changes being noun phrases rather than verb_noun constructions prevent a perfect score.

Tool Count2/5

34 tools is well above the 25-tool threshold for over-scoping, making the set heavy for an agent to navigate. While the server spans many domains, several tools like generate_llms_txt, scan_dependency, and ai_visibility_check feel tangential to the core news/research purpose and would be better split into separate servers.

Completeness4/5

The core news/research domain is well covered: lookup, grounded verification, deep multi-source research, entity profiles, comparisons, subscriptions, and alert feeds are all present with no obvious dead ends. Minor gaps exist—such as no direct full-text article retrieval or a dedicated free-text news search beyond latest_news filters—but agents can work around them.