Skip to main content
Glama

Start Load Test

start_load_test

Start a load test against a target. Max 100 concurrent connections, max 5 min duration, 1 test at a time per target, 60s cooldown between tests. Modes: HTTP_FLOOD (autocannon), BROWSER_USERS (Playwright), COMBINED (both). OWNER/ADMIN only.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
modeNoLoad test modeHTTP_FLOOD
rampUpNoRamp-up time in seconds
durationNoTest duration in seconds
targetIdYesThe target ID to load test
targetUrlNoOverride URL (defaults to target URL)
httpMethodNoHTTP method
concurrencyNoNumber of concurrent connections
acknowledgeSharedOriginNoAcknowledge blast radius if target shares a server with other monitored sites

TDQS

A4.4/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the full burden of behavioral disclosure. It effectively communicates safety limits (100 connections, 5 min duration, 1 test at a time, 60s cooldown), permission requirements (OWNER/ADMIN only), and mode-specific tooling (autocannon, Playwright). It stops short of describing response behavior or error scenarios, but covers critical operational boundaries.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Four dense, purposeful sentences. Each sentence adds distinct information: action, limits, modes, and permissions. No fluff or repetition; front-loaded with the primary purpose.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a complex 8-parameter tool with no output schema, the description covers key operational constraints, permissions, and modes. It doesn't elaborate on each parameter but the schema does. It could mention the response shape or next steps (e.g., get_load_test), but the description is sufficient for starting a load test.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters4/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 100% so baseline is 3. The description adds value by explaining mode enum values with underlying tools and mentioning real-world constraints (cooldown, one-at-a-time) not present in the schema. It also reinforces concurrency and duration limits aligned with schema min/max.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description opens with 'Start a load test against a target', using a specific verb and resource. It clearly distinguishes from sibling tools like cancel_load_test or get_load_test by focusing on the initiation action, and further details modes and constraints.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description provides clear operational context: max concurrency, duration, single test per target, cooldown, modes, and OWNER/ADMIN permission. It does not explicitly state exclusions or compare directly to alternatives like trigger_test, but the limits and constraints give strong usage guidance.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

TDQS

A3.5/5.0
Disambiguation4/5

Most tools are clearly separated by resource (targets, runs, findings, incidents, etc.) and action. A few close pairs like active_runs/list_runs and mute_finding/create_muting_rule could confuse, but descriptions clarify the distinctions.

Naming Consistency4/5

The majority of tools follow verb_noun naming (create_target, get_target, delete_journey). A few outliers use noun phrases (active_runs, daily_trends, system_health, team_stats) which slightly breaks the pattern, but overall the convention is predictable.

Tool Count1/5

74 tools is extreme for any MCP server. Even for a comprehensive monitoring platform, this overwhelms agents with too many granular operations (e.g., enable_all_tests vs disable_all_tests vs update_test, or import_targets duplicating create_target). A more consolidated set would be appropriate.

Completeness5/5

The tool surface is remarkably complete for the monitoring domain: full CRUD for targets, journeys, rules, reports, secrets, and fragments; plus run triggering, incident management, findings handling, SEO tracking, guest scans, and admin tools. Only maintenance windows lack an update operation, which is minor.

Resources