Skip to main content
Glama

Datasets

datasets
Read-onlyIdempotent

Search the Michigan Open Data catalog of open datasets by keyword. Returns each dataset's resource_id, name, description, category and update date — pass the resource_id to query/metadata.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
limitNoMax datasets (1-100, default 20).
queryNoKeyword to search dataset titles/descriptions (e.g. "budget", "crime", "health").
offsetNoPagination offset.

Schema Changelog

Changes observed during successful MCP inspections. Dates show when Glama detected each change.

  1. Changed1 schema field changed
    • addedInput schema / examples
      Added value: +[
      +  {
      +    "query": "budget"
      +  },
      +  {
      +    "limit": 50,
      +    "offset": 0,
      +    "query": "crime"
      +  }
      +]
  2. First observed

TDQS

A4.1/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations already declare readOnlyHint, openWorldHint, idempotentHint, destructiveHint, so the safety profile is covered. The description adds value by specifying return fields and the chaining pattern, which are not in annotations. However, it does not disclose rate limits or pagination behavior beyond schema.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is two sentences, each purposeful. The first defines the action, the second specifies the output and a usage hint. It is front-loaded and contains no redundant or irrelevant information.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given the tool's simplicity (3 parameters, no output schema, strong annotations), the description sufficiently covers purpose, return values, and next steps. No additional context is necessary for correct invocation.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 100% with well-described parameters. The description adds context by explaining the 'query' parameter as a keyword and implies pagination via limit/offset, but does not significantly extend beyond the schema's existing descriptions. Baseline 3 is appropriate.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly states the tool searches the Michigan Open Data catalog by keyword and lists the specific return fields (resource_id, name, description, category, update date). This specific verb-resource combination distinguishes it from sibling tools, which are unrelated (e.g., deep_research, bet_research).

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines3/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description implies usage for searching Michigan open data and hints at chaining with 'pass the resource_id to query/metadata', but it does not explicitly state when not to use it or differentiate from siblings like 'query'. Usage context is clear but exclusions are missing.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

TDQS

A3.8/5.0
Disambiguation2/5

There are multiple severe overlap clusters. ask_pipeworx, ask_pipeworx_beta, ask_pipeworx_grounded, and deep_research all route to the same 5,578 tools, and ask_pipeworx_beta explicitly states it 'currently matches ask_pipeworx exactly' — a direct ambiguity. The six polymarket_* tools plus bet_research form another dense, hard-to-distinguish cluster, and entity_profile/recent_changes/compare_entities/resolve_entity all have overlapping entity-investigation purposes. The long descriptions help but an agent would frequently misselect.

Naming Consistency3/5

The dominant families are internally consistent (polymarket_* prefix, ask_pipeworx_* suffix family, and the verb-based remember/recall/forget), which aids navigation. However, the overall set mixes several conventions: single-word nouns (query, datasets, metadata, recall), verb_noun compounds (validate_claim, generate_llms_txt), and domain_noun names (entity_profile, polymarket_edges). Readable, but there is no unified pattern across the server.

Tool Count2/5

34 tools is clearly over the 25-threshold for heaviness, and several earn little distinct value: ask_pipeworx_beta is a live duplicate of ask_pipeworx, the five-algorithm Polymarket family could be consolidated, and meta/utility tools (suggest_questions, discover_tools, pipeworx_trending, generate_llms_txt, scan_dependency) feel bolted on rather than essential. The breadth of the data domain justifies some size, but the redundancy and tangents push it into bloat.

Completeness3/5

Within its core sub-domains the surface is fairly complete: company research has resolve→profile/compare→recent_changes→validate_claim as a full lifecycle, subscriptions have subscribe/unsubscribe/list/recent_alerts, and memory has remember/recall/forget. The Polymarket workflow is especially thorough (detect→verify→fill-risk→track-decay). However, the server's stated identity ('Data Michigan') is barely served — the Michigan Open Data surface is only search/schema/query with no update or write path — and the scatter of unrelated tools (npm dependency scan, llms.txt generation) makes the overall purpose incoherent, so gaps are hard to evaluate.