Skip to main content
Glama

introspection_system_probe

Read-onlyIdempotent

Return the full evidence trace for a single federation node. Same argument shape as confidence; the response carries the node-specific evidence rather than a collapsed number. Optional repo scopes to one repo (Phase A). Valid node ids come from introspection_system_list_nodes. Returns: The probe result for the requested target. Example: call introspection_system_probe with arguments {}.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
repoNo
node_idYes
node_kindYes

Schema Changelog

Changes observed during successful MCP inspections. Dates show when Glama detected each change.

  1. Changed2 schema fields changed
    • addedInput schema / properties / node_id / maxLength
      Added value: +4000
    • addedInput schema / properties / repo / maxLength
      Added value: +4000
  2. First observed

TDQS

C2.9/5.0
Behavior3/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations already declare the tool as read-only (readOnlyHint=true) and non-destructive (destructiveHint=false), so the description doesn't need to restate that. It adds that the response carries 'node-specific evidence rather than a collapsed number' and that valid node ids come from another tool. However, it doesn't describe the structure of the evidence trace or any error behavior, missing an opportunity to enrich beyond what annotations provide.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness2/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is somewhat verbose with an unhelpful example ('call introspection_system_probe with arguments {}') that implies no required parameters, contradicting the schema. The phrase 'Returns: The probe result for the requested target' is redundant and adds little. Could be tightened to focus on the evidence trace and required parameters.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given no output schemabool, the description should clarify what the 'evidence trace' looks like or at least mention that it differs from a scalar confidence value. It references introspection_system_list_nodes for valid ids, which is helpful, but lacks details on node_kind values, error conditions, and the overall behavior for different node types. The misleading example further reduces completeness.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters2/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 0% (description only references 'repo' and the argument shape relative to 'confidence'). The required parameters node_kind and node_id are not described beyond the example call, which is incorrect (empty {} ). The mention of 'repo' being optional adds some value, but the enum values for node_kind are left to the schema without explanation.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description states a specific action ('Return the full evidence trace') with a clear resource ('a single federation node'). It hints at differentiation from the 'collapsed number' tool (likely confidence) but does not explicitly name or contrast with the sibling introspection_system_confidence. The reference to 'same argument shape as confidence' provides context but could be clearer about the distinguishing purpose.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines3/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Provides some usage guidance: mentions that valid node ids come from introspection_system_list_nodes Light, and that `repo` optionally scopes to a single repo. However, it does not explicitly state when to use this tool over alternatives (like confidence or list_nodes), nor does it mention the requirement for node_kind and node_id as mandatory parameters beyond the example with empty arguments, which is misleading.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

TDQS

A3.6/5.0
Disambiguation4/5

Tools are grouped by clear domain prefixes (federation_*, introspection_*, moltbook_*) with each targeting a distinct resource+action (bond_post/release/status, key_challenge/bind/status, journal_append/read). The primary near-overlap — catalog_search_multi vs catalog_search_grouped_multi — is explicitly disambiguated in descriptions. Minor confusion risk exists among the four knowledge tools (federation_help, federation_why, about_us_about, how_to_about) but their purposes (how/why/manual/walkthrough) are distinct enough.

Naming Consistency3/5

The dominant `federation_<verb>_<noun>` pattern (create_tenant, list_agents, bond_release) is strong, but it's mixed with bare-noun tools (federation_arena, federation_offer, federation_solvency, federation_pricesheet, federation_help) and noun-noun variants (federation_manager_tree, federation_tenant_info). Non-federation tools use a loose `<domain>_<verb>` or single-token convention (legal_get, web_research, about_us_about). Readable overall, but conventions are noticeably mixed across the surface.

Tool Count2/5

At 72 tools this crosses the 50+ threshold for an extreme count. While the federation's scope is genuinely broad (manager lifecycle, tenants, catalog, agents, bonds, keys, journal, canon, introspection, social, email, research), the surface is bloated — roughly 15 bare introspection tools (list_nodes, probe, confidence, diff, coverage_gaps, co_decisions, climb_history, change_graph, change_reach, corpus_*) cover meta-self-knowledge that could plausibly collapse into fewer verbs. Agents would face a very large selection space.

Completeness5/5

The tool surface is exhaustively complete for the federation domain: applicant and operator sides of admittance, full manager lifecycle (create/list/freeze/attest/bond/key), full tenant lifecycle (create/list/info/update/suspend/delete/enter), catalog discovery with change-detection, pricing, solvency, latency, journaling, governance, legal, and even external outreach (email/moltbook/research). No dead ends exist — every write has a corresponding read/status path, and branched platform tools are intentionally deferred behind enter_tenant rather than omitted.

Resources