Skip to main content
Glama
bill-kopp-ai-dev

Claude Code CLI MCP Server

claude_start_task

Start long-running Claude Code tasks (up to 3600s) in the background and get a run ID immediately. Monitor progress with poll, stop with cancel.

Instructions

Start a Claude Code CLI task asynchronously (up to 3600s).

Returns immediately with a run_id. The subprocess continues running in the background; use claude_poll_task to monitor and claude_cancel_task to stop. Use this for any task that may exceed 600s (architecture, migration, large multi-file refactors, long-running data agents).

Required: workspace_path, prompt. Optional: model, fallback_model, permission_mode, options.timeout_s (default 300, max 3600 for async), capture_changes.

Concurrent run limit: bounded by Settings.max_concurrent_runs. New calls beyond that limit return MAX_CONCURRENT_RUNS_EXCEEDED.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
reqNo

Output Schema

TableJSON Schema
NameRequiredDescriptionDefault
run_idYes
session_idYes
started_atYes

Schema Changelog

Changes observed during successful MCP inspections.

  1. First observedv0.2.0

TDQS

A4/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations provided, the description carries the full burden, and it does well: async background execution, immediate return, timeout boundaries, and the MAX_CONCURRENT_RUNS_EXCEEDED error are all disclosed. It could go further on side effects or cleanup expectations, but the core behavioral traits are clearly stated.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is tightly written and well structured, with the most important operational facts front-loaded. Each paragraph earns its place: async behavior, return semantics, when to use it, required parameters, and concurrency limits.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness3/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

The description covers the main async workflow and monitoring guidance well, but the tool is complex with many optional parameters and no annotations. It omits the likely synchronous alternative claude_run_task and does not clarify how to pass required fields inside the 'req' object, which is a meaningful invocation gap.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters2/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 0%, so the description must compensate, but it only names a subset of parameters and does not explain the required 'req' object wrapper. More importantly, it states options.timeout_s default is 300 while the schema says 600, creating a direct contradiction that could lead to incorrect invocation.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description opens with a specific verb and resource: 'Start a Claude Code CLI task asynchronously'. It clearly distinguishes itself from siblings by emphasizing async execution, immediate return of a run_id, and the ability to run up to 3600s.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description gives explicit guidance for when to use this tool: 'Use this for any task that may exceed 600s' with concrete examples. It also directs the agent to claude_poll_task and claude_cancel_task for lifecycle management. However, it does not explicitly contrast with claude_run_task, which appears to be the likely synchronous alternative.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.