Skip to main content
Glama
wirkspace

wirk-mcp

Official
by wirkspace

wirk-mcp

wirk-mcp gives agents WIRK over MCP: five tools, wirk_status, wirk_query, wirk_write, wirk_review and wirk_show, whose arguments are WIRK's request bodies and whose answers are WIRK's compact text. Coordination and ticketing built for agents and the people they work with: your wirk, its context and its evidence in one wirkspace.

Install

One command installs the wirk command and this server, sets up Claude Code and Codex when they are present, and logs in:

curl -fsSL https://wirk.life/install | sh

By hand: uv tool install wirk-mcp (or the wheels attached to a release), wirk login, then point your agent host at the installed binary:

claude mcp add --scope user wirk -- "$(command -v wirk-mcp)"      # Claude Code
codex mcp add wirk -- "$(command -v wirk-mcp)"                    # Codex

The server uses the configuration wirk login made (~/.config/wirk, or $WIRK_CONFIG_DIR) and only the agents' token in it; it never reads a person's own token. A new login needs no restart.

Related MCP server: MCP Customer Support AI

Use

An agent starts with wirk_status, optionally with a one-line task. wirk_query fetches by ID, short ID or exact title, lists with fields, finds what matters for the words in about, or looks up a receipt. wirk_write changes items and links in one batch, stating in expect the revision it read (rN on a card); completing wirk gives its evidence as reason. Context changes apply when the agent may make them; otherwise they are refused with requires_review and the agent proposes them. Only people decide proposals, so an agent's wirk_review is refused (not_authorized for its own proposal, person_required for any other); the person decides at their own terminal with wirk review ID@N accept --reason WHY --person (see the wirk command). wirk_show makes a page a person can open; it is not live on api.wirk.life yet and answers views_unavailable. format: "json" returns data instead of text. request_id may be left out: the server makes one and names it, and resending the identical body with it applies once.

What leaves the machine

The requests the agent makes and their content; the token, only as the bearer header to the configured address; and with wirk_status, the host's name and version, a session number for this server process, and the repository (host/owner/name, or a hash) and branch of the directory it runs in. No telemetry.

Development

The tests need the wirk package installed beside this one: uv pip install -e ../wirk-cli && uv pip install -e . && python -m pytest tests -q.

Releasing

A v* tag matching pyproject.toml's version tests against the wirk release of the same version, attaches the wheel, the sdist and SHA256SUMS to a GitHub Release, and publishes to PyPI through trusted publishing once the wirk-mcp project trusts this repository's release.yml in the pypi environment and the repository variable PYPI_PUBLISH is true. No token is stored anywhere.

License

Apache-2.0. See LICENSE.

Available Tools

5 tools
wirk_queryC

Fetch by ID, short ID or exact title; list with fields; find what matters for the words in about, ranked by meaning; or look up a receipt.

ParametersJSON Schema
NameRequiredDescriptionDefault
sortNo
aboutNoWords; items ranked by meaning, or containing every word when ranking is unavailable.
depthNo
fetchNoRefs: ID, short ID, exact title, or ID@N for revision N.
limitNo
cursorNo
fieldsNoFilters, e.g. {"status": ["open", "in_progress"], "kind": "work", "owner": "me", "text": "exact words", "linked": "ID", "changed_days": 7}; kind is work, context, folder or doc; state is an initiative's; proposal=["proposed"] lists proposals.
formatNo
receiptNoA request_id: its stored write or review receipt.
max_bytesNo
workspace_idNo

TDQS

C2.8/5.0
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

No annotations are provided, so the description carries the full behavioral burden for an 11-parameter tool. It discloses one useful trait — fallback to 'containing every word when ranking is unavailable' — but says nothing about pagination (limit/cursor), byte capping (max_bytes), depth of returned data, or whether the operation is read-only.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness3/5

Is the description appropriately sized, front-loaded, and free of redundancy?

It is a single compact sentence with no filler, which is efficient, but the telegraphic semicolon-clause style is dense and hard to parse on first read, and the mode-to-parameter mapping is left implicit.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a zero-required, eleven-parameter multi-mode query tool with no annotations and no output schema, the description is too thin. Key behaviors — paging, result shaping via depth/format, and the meaning of workspace_id — are entirely undocumented.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters2/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is only 36%, and the description compensates for just three of the eleven parameters (fetch, about, receipt) — all of which the schema already documents. Sort, depth, limit, cursor, format, max_bytes, and workspace_id remain unexplained anywhere.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description enumerates four concrete capabilities — fetch by ref (ID/short ID/exact title), list with fields, semantic search over 'about' words, and receipt lookup — so the agent knows exactly what the tool can do. It never names or differentiates itself from siblings like wirk_show or wirk_status, which likely overlap on the fetch behavior.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines3/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Usage is only implied through the mapping of verbs to parameters (fetch → 'fetch', list → 'fields', search → 'about', lookup → 'receipt'). There is no explicit when-to-use, when-not-to-use, or routing away from wirk_show/wirk_status despite apparent overlap.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

wirk_reviewA

Only people decide proposals: as an agent this is refused (not_authorized for your own proposal, person_required for any other), and the proposal waits. Your person decides at their terminal: wirk review ID@N accept --reason WHY --person.

ParametersJSON Schema
NameRequiredDescriptionDefault
formatNo
decisionsYes
request_idNoOptional; generated and named in the answer.
workspace_idNo

TDQS

A3.6/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the full behavioral burden and does most of it: it discloses that the call is refused, names the two distinct error codes and which case triggers each, and notes the proposal 'waits' (i.e., no side effect occurs). It doesn't describe the successful path in detail, but since the caller can never reach it that omission is minor.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Compact and front-loaded: the decisive constraint ('as an agent this is refused') appears in the first clause. Every element — error codes, wait behavior, the terminal command — adds information. The CLI fragment is slightly cryptic but earns its place by showing the expected decision shape.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given no annotations, no output schema, and a tool the agent cannot successfully call, the description supplies what the agent actually needs: that it will fail, why, and who should do it instead. The remaining gap is the decisions payload format for a human-driven call, which the CLI example only partly conveys.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is only 25%, and the nested decisions objects (id, revision, action, reason) are undocumented in the schema. The description partially compensates by mapping the CLI 'ID@N' to id+revision and '--reason WHY' to reason, and by exemplifying the 'accept' action. It never explains the decisions array, the accept/reject/defer enum, or the request_id/workspace_id params, so the compensation is only partial.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose3/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description implies the tool submits decisions on proposals ('Only people decide proposals', the CLI form 'wirk review ID@N accept'), but never directly states the verb+resource the agent would be invoking. Instead it foregrounds the refusal, so an agent gets behavior more clearly than the core action. It does loosely separate itself from siblings like wirk_write and wirk_query, but not explicitly.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

It gives strong when-not guidance: as an agent this call is refused, both for your own proposal (not_authorized) and any other (person_required). It then routes to the right alternative — 'Your person decides at their terminal' — making the handoff unambiguous. It stops short of stating any condition under which the agent should call it, which is consistent but keeps it from a 5.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

wirk_showC

Not live yet on api.wirk.life: answers views_unavailable. Make a live, read-only page a person can open, from the preset status or a view config; returns its link. Anyone with the link can open it until it expires.

ParametersJSON Schema
NameRequiredDescriptionDefault
titleNo
blocksNoView blocks; a refusal names the path and the allowed keys.
formatNo
presetNo
revokeNo
workspace_idNo

TDQS

C2.8/5.0
Behavior3/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the full behavioral burden. It does disclose meaningful traits beyond the schema: read-only, link-based access for anyone, and expiry. However, it omits auth/permission requirements, revocation behavior, and any rate/visibility caveats. The ambiguous 'Not live yet' clause further weakens confidence.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness3/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The definition is short and the actionable sentence comes second. The opening 'Not live yet on api.wirk.life: answers views_unavailable' is cryptic and consumes the front-loaded position without conveying a clear condition. Otherwise reasonably efficient.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a 6-parameter tool with no annotations, no output schema, and 17% schema coverage, the description does not do enough. It leaves revoke, format, title, and workspace_id unexplained and gives no return-value detail beyond 'its link', while the 'not live' statement leaves the agent unsure whether the tool functions at all.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters2/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is only 17% across 6 parameters, so the description must compensate and largely does not. It gestures at 'preset status' and 'a view config' (roughly preset/blocks) but gives no meaning for title, format, revoke, or workspace_id. A low-coverage schema is left under-explained.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

The core sentence gives a specific verb and resource: 'Make a live, read-only page a person can open... returns its link.' That is clear enough to distinguish from sibling read/write tools. The leading clause 'Not live yet on api.wirk.life: answers views_unavailable' muddies the message and makes the tool's availability ambiguous rather than clarifying its purpose.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines2/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

No when-to-use guidance and no comparison against siblings (wirk_status, wirk_query, wirk_write, wirk_review). The mention of 'the preset status or a view config' is parameter-level input detail, not usage guidance. The agent gets no signal about when to pick this over an alternative.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

wirk_statusA

Start here. Who you work for, your wirk, what is in progress, what needs your review, recent changes and how to ask for more. Writes nothing.

ParametersJSON Schema
NameRequiredDescriptionDefault
taskNoOptional one line about your task; the items that matter for it are listed.
formatNotext (default) or json.
max_bytesNo
workspace_idNoWirkspace ID, full or short; omit for your only one.

TDQS

A3.6/5.0
Behavior3/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Without annotations, the description carries the full behavioral burden. It discloses 'Writes nothing' which is a key read-only trait, but it omits other relevant behaviors such as permission requirements, response size limits (implied by max_bytes), or rate limits. A score of 3 reflects some useful disclosure but incomplete transparency.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is a short, front-loaded fragment list beginning with 'Start here.' It is concise and avoids waste, though the single run-on sentence could be better structured for scannability. Every phrase earns its place by conveying purpose and scope.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness3/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given no output schema and no annotations, the description covers purpose and read-only behavior reasonably well, but it does not explain the response format, the effect of max_bytes, or how workspace_id and format interact. The 75% schema coverage fills some gaps, but the description leaves a quarter of the parameter surface undocumented.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 75%, so most parameters are documented in the schema. The description adds no parameter-level meaning beyond what the schema provides, and the one undocumented parameter (max_bytes) is not compensated for. Baseline 3 is appropriate when the schema does the heavy lifting.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description states the tool provides a status overview: who you work for, your work, what's in progress, what needs review, and recent changes. It is specific enough to identify the resource, and 'Writes nothing' hints at read-only. However, it does not name or contrast with sibling tools like wirk_query or wirk_show, leaving differentiation to inference.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The opening 'Start here' gives clear contextual guidance for when to use this tool as the entry point. It does not explicitly name alternatives or state exclusions, but the implied first-step role is strong. No explicit when-not guidance is present.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

wirk_writeA

Change items and links in one atomic batch. Put the revision you read (rN) of each existing item you change or link from in expect; one left out is refused as basis_changed. Context changes apply when you may make them; otherwise they are refused with requires_review: propose them (mode propose, with a reason). After an uncertain result, resend the identical body with the request_id it names.

ParametersJSON Schema
NameRequiredDescriptionDefault
modeNo
expectNo{"<ID>": <revision you read>}, e.g. {"5c1e7a90": 3}.
formatNo
reasonNoWhy: required to propose and to archive; to complete wirk, the evidence (tests, link, file path or upload ID).
operationsYes1-32 of: {op:item.create, ref?, data:{title, body?, work?:{owner_id?, due_at?, criteria?:[{text}]}, context?:{level, state?, steward_id?, open?}, fields?}, allow_duplicate_of?} · {op:item.edit, id, patch:{title?, body?, work?, fields?}} · {op:link.create, data:{type, from, to}}, type related_to, contributes_to, requires, cites or relies_on, from/to may be $ref of a new item; a cites quotation must match the cited text exactly, or quotation_mismatch · {op:link.remove, id} · {op:item.archive, id} · {op:item.restore, id}. work makes an item work, context makes it context (level organization or initiative); fields maps a field to an option, null clears.
request_idNoOptional; generated and named in the receipt. After an uncertain result, resend with that ID.
workspace_idNo

TDQS

A3.5/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the full burden and does well: it discloses atomicity, optimistic concurrency via revisions (basis_changed refusal), a requires_review refusal path, and idempotent retry via request_id after an uncertain result. It omits what the operation returns and any permission/rate-limit context, but the failure-mode disclosure is rich.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness3/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Three dense sentences, front-loaded with the core action, but the phrasing is cryptic ('Context changes apply when you may make them') and jams multiple error codes, modes, and retry instructions into a single line. Efficient in size, weak in readability.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness3/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a 7-parameter tool with nested objects and no output schema, the description covers behavior well but leaves real gaps: workspace_id and format are unexplained, and with no output schema the 'receipt' the description references is never described. Adequate but not complete.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 57%, and much of the description's parameter text (expect format, reason, request_id) restates the schema's own descriptions rather than extending them. It does reinforce the expect-as-basis semantic and the request_id retry contract, but adds little syntax or meaning beyond the schema.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a specific verb and resource: 'Change items and links in one atomic batch.' The atomic-batch qualifier is meaningful and tells the agent this is a multi-operation write. It does not distinguish itself from the read-oriented siblings (wirk_query, wirk_show, wirk_status), but the write/read split is inferable.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines3/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Gives conditional guidance for mode propose ('Context changes apply when you may make them; otherwise they are refused with requires_review: propose them') and for retries, which is genuinely useful routing. However it never compares against sibling tools or states when NOT to use this tool, so usage is implied rather than explicit.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Tool Schema Changelog

Recent tool additions, removals, and schema changes observed during successful MCP inspections.

  1. 5 tool updatesv0.4.0
    • First observedwirk_query
    • First observedwirk_review
    • First observedwirk_show
    • First observedwirk_status
    • First observedwirk_write

TDQS

A3.5/5.0

Scored across 5 tools

Disambiguation4/5

Each tool has a fairly distinct role: status is a read-only overview, query is targeted fetch/list/search, write mutates, review handles governance, show publishes a page. status and query both read, which creates some mild overlap, but the 'start here' framing and query's specificity keep them separable.

Naming Consistency5/5

All tools follow a clean wirk_<verb> snake_case pattern with no style mixing. The only minor quirk is that 'status' is a noun while the rest are verbs, but the prefix-and-shape convention is fully consistent.

Tool Count5/5

Five tools is well-scoped for a work-coordination server, with each mapping to a distinct lifecycle stage (overview, read, mutate, review, publish). Nothing feels redundant or missing at the count level.

Completeness4/5

Read (status/query), mutate (write handles create/update/link atomically), and governance (review) are all covered, giving a coherent lifecycle. The notable gap is wirk_show, which is admitted to be non-functional ('views_unavailable'), leaving a dead end in the surface.

Maintenance

ActivityMaintained
ResponsivenessNo issues

Related MCP Connectors

Related MCP Servers

  • A
    license
    A
    quality
    D
    maintenance
    Enables AI models to perform IT service management operations on Freshservice, including managing tickets, changes, problems, releases, assets, projects, and more through a set of MCP tools.
    36
    2
    MIT
  • F
    license
    Not graded
    quality
    C
    maintenance
    Enables AI clients to access and manage an internal support ticket queue through MCP tools, resources, and prompts, including searching and viewing tickets, adding comments, closing tickets with confirmation, reading knowledge base articles, and viewing queue summaries over OAuth-secured Streamable HTTP.
    -