Skip to main content
Glama

List World Model Evaluation Runs

lyzr_world_model_list_evaluation_runs
Read-onlyIdempotent

Retrieve evaluation runs for a World Model by ID. Review performance and track model evaluations.

Instructions

List all evaluation runs created for a given World Model.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
world_model_idYesThe World Model id whose evaluation runs to list
Behavior3/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations already declare readOnlyHint, openWorldHint, and idempotentHint, covering the safety profile. The description adds the scope constraint 'for a given World Model' but that is also present in the schema. No additional behavioral details like pagination or ordering are disclosed.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is a single, front-loaded sentence with no filler. It directly states the action and scope with minimal words.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a simple 1-parameter list tool with strong annotations, the description is sufficient. It clearly states what is listed and the required scope. It does not explain the return structure, but no output schema is present and the purpose is straightforward.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 100% and the parameter world_model_id has a clear description in the schema. The tool description does not add any extra meaning beyond what the schema already provides, so a baseline 3 is appropriate.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description uses the specific verb 'List' and clearly identifies the resource ('evaluation runs') and the scope ('for a given World Model'). This distinguishes it from sibling tools like get_evaluation_run and create_evaluation_run.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines3/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description implies that this tool is used to list all runs for a world model, but it does not explicitly mention when to use it versus alternatives like get_evaluation_run. No when-not-to-use guidance is provided.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Install Server

Other Tools

Latest Blog Posts

MCP directory API

We provide all the information about MCP servers via our MCP API.

curl -X GET 'https://glama.ai/api/mcp/v1/servers/nandanNM/lyzr-mcp'

If you have feedback or need assistance with the MCP directory API, please join our Discord server