Skip to main content
Glama
malkreide

swiss-efv-mcp

by malkreide

πŸ‡¨πŸ‡­ Part of the Swiss Public Data MCP Portfolio β€” open-source MCP servers connecting AI agents to Swiss public and open data. This is a private project. It is independent of any employer or institutional affiliation.

πŸ›οΈ swiss-efv-mcp

Version CI License: MIT Python 3.11+ MCP Auth: none Portfolio

MCP server for Swiss federal finances (EFV): budget, debt, forecasts and spending by task and institution.

πŸ‡©πŸ‡ͺ Deutsche Version

Overview

This server closes the fiscal gap in the portfolio's Economics & Finance cluster. swiss-snb-mcp already covers monetary policy; swiss-efv-mcp adds the state budget β€” federal revenue, expenditure, balance, debt ratios (with forecasts to 2029), a hierarchical budget drill-down, and spending by department. Data comes from the EidgenΓΆssische Finanzverwaltung (EFV) via opendata.swiss (OGD Schweiz).

Related MCP server: ch-eli-mcp

Features

  • Five read-only tools over the curated EFV FS/GFS dump files.

  • Headline series 1990–2029 per household (bund, ktn, gdn, staat, sv) and model (FS / GFS); every point carries is_projection so actuals and plan/forecast years are unambiguous.

  • Hierarchical federal-budget drill-down and spending by department / unit.

  • 24 h TTL in-memory cache with stale-serve fallback; retry with exponential backoff (2/4/8 s); dump_status never returns empty silently.

  • Dual transport: stdio (local) and SSE (cloud).

  • No authentication required β€” public open-government data (No-Auth-First).

🎯 Anchor Demo Query

"How has the federal balance developed since the SNB rate turnaround in 2022 β€” and which task areas absorbed the growth in spending?"

fiscal_headline(variable="saldo", household="bund", year_from=2021)
fiscal_budget_breakdown(topic="Ausgaben nach Aufgabengebiet", level=2)

Cross-read with swiss-snb-mcp, this connects the interest-rate cycle to the federal deficit β€” something neither server can answer alone.

Demo

Demo: Claude using fiscal_headline and fiscal_budget_breakdown

Prerequisites

  • Python 3.11+

  • uv / uvx (recommended) or pip

  • Network access to data.finance.admin.ch and efv.admin.ch β€” no API key needed

Installation

uvx swiss-efv-mcp            # zero-install run (once published to PyPI)
# or
pip install swiss-efv-mcp

Claude Desktop (claude_desktop_config.json):

{
  "mcpServers": {
    "swiss-efv": {
      "command": "uvx",
      "args": ["swiss-efv-mcp"]
    }
  }
}

Quickstart

# Run locally over stdio (default transport)
uvx swiss-efv-mcp

# From a checkout, without installing
PYTHONPATH=src python -m swiss_efv_mcp

Configuration

All configuration is loaded once into a typed Settings object (pydantic-settings). The legacy unprefixed names below keep working; the canonical names use the EFV_MCP_ prefix. Defaults are safe for local use.

Variable

Default

Purpose

TRANSPORT

stdio

Transport: stdio (Claude Desktop) or sse / streamable-http (cloud)

HOST

127.0.0.1

Bind host (SSE only). Loopback by default; set 0.0.0.0 only in a container

PORT

8000

Bind port (SSE only)

EFV_MCP_LOG_LEVEL

INFO

structlog level (JSON to stderr)

EFV_MCP_CORS_ORIGINS

[]

SSE only: explicit allowed browser origins (default-deny; comma-separated or JSON)

EFV_MCP_OTEL_ENABLED

false

Enable OpenTelemetry tracing (requires the otel extra); standard OTEL_* env vars configure export

Cloud (Render / Railway):

TRANSPORT=sse PORT=8000 swiss-efv-mcp   # exposes /sse

Available Tools

Tool

Purpose

fiscal_headline

Revenue / expenditure / balance / debt ratios over 1990–2029, per household and model; every point flags is_projection

fiscal_budget_breakdown

Hierarchical federal budget by topic (Ausgaben nach Art / nach Aufgabengebiet, Einnahmen, Bilanz, …)

fiscal_by_institution

Spending per department / administrative unit since 2007 (Personalausgaben, Informatik, external services, FTE)

fiscal_list_dimensions

Discover valid parameter values β€” call this first to build correct arguments

fiscal_status

Cache freshness and upstream health per dataset; never returns empty silently

dump_status

Deprecated alias of fiscal_status (kept for backward compatibility; removed in a future minor)

All tools are read-only: each is annotated readOnlyHint: true, destructiveHint: false, only issues HTTP GETs against the EFV dump files, and has no write, send, or filesystem capability.

MCP primitives. This server uses only the Tools primitive. The EFV data are sliced live from cached dumps with no stable resource hierarchy to expose as Resources, and there are no server-authored Prompts. The five tools are small and closely related, so they live in a single server.py rather than a tools/ package.

Architecture

                      β”Œβ”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”
   Claude / Agent ──▢ β”‚  swiss-efv-mcp (FastMCP)      β”‚
                      β”‚  5 tools Β· Pydantic v2 env.   β”‚
                      β””β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”¬β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”˜
                                      β”‚ fetch + retry + TTL cache
              β”Œβ”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”΄β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”
              β–Ό                                               β–Ό
   data.finance.admin.ch                          efv.admin.ch/dam
   fs_dashboard/main_extern.csv                   bundeshaushalt_de.csv
   (headline, 1990–2029)                          institutionen_de.csv

Architecture decision

This server uses Architecture C (Dump-first).

Rationale (verified live on 2026-07-24):

  • The EFV FS/GFS dashboard has no filtered query API; it serves static CSV dumps that its front-end filters in the browser.

  • Three curated files are small enough to fetch-and-cache whole (516 KB / 5 MB / 1 MB). They cover the headline aggregates, the hierarchical budget and the by-institution view β€” i.e. the answerable questions.

  • The full detail cubes (standardauswertung.csv 157 MB, fir_art_funk.csv 1.23 GB) are out of scope for v0.1.0; loading them per request is not viable. A future Phase 2 would pre-process them into SQLite/Parquet.

Consequences:

  • Files are cached in memory with a 24 h TTL; stale cache is preferred over an empty response when upstream is down.

  • Retry with exponential backoff on all HTTP; dump_status always returns a readable state.

Project Structure

swiss-efv-mcp/
β”œβ”€β”€ src/swiss_efv_mcp/
β”‚   β”œβ”€β”€ __init__.py
β”‚   β”œβ”€β”€ __main__.py        # entry point; dual transport (stdio / SSE+CORS)
β”‚   β”œβ”€β”€ client.py          # dump-first data layer: egress allow-list, retry, UA, TTL cache
β”‚   β”œβ”€β”€ logging_config.py  # structlog JSON to stderr
β”‚   β”œβ”€β”€ models.py          # Pydantic v2 envelopes (source + provenance)
β”‚   β”œβ”€β”€ server.py          # 5 FastMCP tools (annotated) + testable *_impl functions
β”‚   └── settings.py        # typed pydantic-settings config
β”œβ”€β”€ tests/                 # respx mock tests + hardening tests + @pytest.mark.live
β”œβ”€β”€ docs/                  # network-egress.md + accepted-risk ADRs
β”œβ”€β”€ audits/                # MCP best-practice audit runs (findings, report, summary)
β”œβ”€β”€ README.md Β· README.de.md Β· CHANGELOG.md Β· SECURITY.md Β· CONTRIBUTING.md
β”œβ”€β”€ Dockerfile Β· server.json Β· LICENSE
└── pyproject.toml

Safety & Limits

  • Read-only. Every tool is annotated readOnlyHint: true, only issues HTTP GETs against the EFV dump files, and has no write, send, or filesystem capability.

  • Egress allow-list. An immutable ALLOWED_HOSTS frozenset + assert_host_allowed() is enforced before every request (HTTPS-only, two fixed EFV hosts). URLs are hardcoded constants; no user input builds a URL. See docs/network-egress.md.

  • TLS on. httpx certificate verification is on by default and never disabled.

  • No credentials. The endpoints are public OGD; no API keys or secrets are stored or forwarded. A browser User-Agent is injected because the endpoints 403 the default httpx/curl UA (see Known limitations) β€” do not remove it.

  • Error masking. mask_error_details=True plus client-side masking keep raw upstream/internal detail out of tool results; full detail goes only to the structlog stderr log.

  • Input bounds. Tool arguments carry explicit Pydantic constraints (year 1900–2100, level 1–8, string max_length).

  • Graceful degradation. Retry with exponential backoff (2/4/8 s); a stale cache is served over an empty response; dump_status always returns a readable state and never a silent empty.

  • Loopback + default-deny CORS. SSE binds to HOST, default 127.0.0.1; set HOST=0.0.0.0 only inside a container (the provided Dockerfile does). Browser origins must be listed explicitly via EFV_MCP_CORS_ORIGINS.

  • Audited. Reviewed against the portfolio MCP best-practice catalogue (44 applicable checks) β€” see audits/ and SECURITY.md. Accepted risks are documented as ADRs under docs/adr/.

  • Not authoritative. Figures are not official; consult the EFV originals for official use.

Known limitations

Live-probe findings (2026-07-24), also in CHANGELOG.md β†’ Known findings:

Finding

Impact

Endpoints return HTTP 403 without a browser User-Agent

UA is injected by the client; do not remove it

opendata.swiss "CSV" links for 2 datasets point to an HTML landing page

real files resolved to a DAM path (/dam/de/sd-web/{id}/…) whose opaque id may rotate on re-upload

NA appears as a literal string in hh/model/source

cleaned to None centrally

"Forward-looking" is not one label: Bund uses "Budget/financial plans", staat uses "Forecasts"

abstracted via is_projection

Accounting-model break at 2022/2023 ("bis 2022" vs "ab 2023" topics)

series has a seam; a note flags affected topics

Detail cubes (157 MB / 1.23 GB) not served

Phase 2; use the curated files for now

Project Phase

This server is in Phase 1 (read-only). Every tool only ever fetches the public EFV dump files β€” there are no write, send, or filesystem capabilities.

Phase

Scope

Status

1 β€” Read-only

Headline series, budget breakdown, spending by institution

βœ… current

2 β€” Detail cubes

Pre-process the 157 MB / 1.23 GB cubes to SQLite/Parquet

planned

3 β€” Multi-agent

(none planned)

β€”

A transition to a later phase would require a re-audit before any write-capable tool is added.

MCP Protocol Version

This server is native to MCP spec 2026-07-28 and still serves the older handshake era, so both pins are stated β€” a single number would describe only half of what clients actually get.

Era

Revision

How a connection negotiates it

Pinned as

modern (default)

2026-07-28

server/discover + a per-request envelope; no initialize handshake, and Client.initialize_result is None

MCP_MODERN_PROTOCOL_VERSION

handshake (legacy clients)

2025-11-25

the classic initialize handshake

MCP_HANDSHAKE_PROTOCOL_VERSION

Both constants live in server.py and are held against the mcp SDK's own LATEST_MODERN_VERSION / LATEST_HANDSHAKE_VERSION rather than against copied-out spec text, and both eras are exercised over a real connection β€” a protocol-changing SDK bump fails CI loudly instead of drifting silently (ARCH-012).

The 2026-07-28 era carries consequences beyond the number:

  • Routing headers. Every modern request carries Mcp-Protocol-Version, Mcp-Method and (for tools/call) Mcp-Name. They are listed in the CORS allow-list in __main__.py; without them a browser client fails at the preflight and never reaches the server. Mcp-Param-* is deliberately absent β€” no tool schema here carries the x-mcp-header annotation that would make a client send one, and a test fails the day one does.

  • Logging is deprecated (SEP-2577). Tool handlers no longer send client-facing log notifications; per-call diagnostics go to the structlog stderr stream, honouring EFV_MCP_LOG_LEVEL. Progress reporting is unaffected and stays.

  • fastmcp>=4.0 is a floor, not cosmetics. Only fastmcp 4 pulls in mcp 2.x, and only there does revision 2026-07-28 exist at all. Under fastmcp 3.x this server would speak 2025-11-25 at best.

Dependencies are kept current via monthly Dependabot PRs (.github/dependabot.yml); protocol-relevant bumps are noted in CHANGELOG.md.

Testing

PYTHONPATH=src pytest tests/ -m "not live"   # offline, respx-mocked
PYTHONPATH=src pytest tests/ -m live         # hits the real EFV endpoints
PYTHONPATH=src ruff check src tests

Changelog

See CHANGELOG.md.

Contributing

Issues and pull requests are welcome. Please keep tools read-only, run ruff check and the offline test suite before submitting, and add a CHANGELOG.md entry under [Unreleased] for user-facing changes. See CONTRIBUTING.md.

Maintainers: see PUBLISHING.md for the step-by-step PyPI release process (Trusted Publishing via GitHub Release).

Security

See SECURITY.md for the security posture, hardening controls, and how to report a vulnerability.

License

MIT for this server β€” see LICENSE. The EFV data remain subject to the OGD Schweiz terms (freely usable, with attribution).

Author

Hayal Oezkan Β· github.com/malkreide

  • Data: EidgenΓΆssische Finanzverwaltung EFV via opendata.swiss (OGD Schweiz, freely usable)

  • Companion: swiss-snb-mcp (monetary policy) β€” the fiscal/monetary pair

  • Portfolio index: swiss-public-data-mcp

Disclaimer: private project, independent of any employer or institution. No warranty; figures are not authoritative β€” consult the EFV originals for official use.

Available Tools

6 tools
dump_statusDump StatusA
Read-onlyIdempotent

DEPRECATED β€” use fiscal_status. Kept as an alias for backward compatibility; will be removed in a future minor release.

Reports cache freshness and upstream health per dataset (SEC-022: every tool now shares the fiscal_ server-identity namespace).

ParametersJSON Schema
NameRequiredDescriptionDefault

No parameters

Output Schema

ParametersJSON Schema
NameRequiredDescription
sourceNo
healthyYes
messageYes
datasetsYes

TDQS

A4.6/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations already declare readOnlyHint, idempotentHint, and non-destructive behavior. The description adds useful behavioral context beyond that: alias behavior, deprecation status, removal timeline, and the SEC-022 server-identity namespace note. No contradiction exists.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is compact and front-loaded with the deprecation warning. The SEC-022 parenthetical is mildly tangential, but the overall structure is efficient and easy to scan.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given zero parameters, an output schema, and a clear alias relationship to fiscal_status, the description covers everything an agent needs: what the tool reports, that it is deprecated, and where to route usage instead.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters4/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

The tool has zero parameters and an empty schema, so the baseline is 4. The description adds no parameter-specific detail, but none is needed since the schema fully documents that the tool takes no inputs.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly identifies the tool as a deprecated alias and states its function: 'Reports cache freshness and upstream health per dataset.' It also names the replacement, fiscal_status, which distinguishes it from the sibling tools.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines5/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

It explicitly says 'DEPRECATED β€” use fiscal_status' and warns it will be removed in a future minor release. This gives unambiguous guidance on when not to use this tool and which alternative to select.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

fiscal_budget_breakdownFiscal Budget BreakdownA
Read-onlyIdempotent

Hierarchical federal-budget breakdown for one topic and year.

Use case: see where the money goes β€” e.g. "which task areas absorbed the spending growth?". topic e.g. 'Ausgaben nach Aufgabengebiet', 'Ausgaben nach Art', 'Einnahmen'. level is the hierarchy depth (1 = total, 2 = first breakdown …); 'contains' filters the path substring for drill-down. An empty result carries a note suggesting a different level or topic.

ParametersJSON Schema
NameRequiredDescriptionDefault
yearNo
levelNo
topicNoAusgaben nach Aufgabengebiet
containsNo

Output Schema

ParametersJSON Schema
NameRequiredDescription
noteNo
yearYes
itemsYes
levelYes
topicYes
sourceNo
provenanceYesdump = freshly fetched CSV, cached = in-memory

TDQS

A4.6/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

The annotations already cover read-only, idempotent, and non-destructive behavior, so the description adds value by explaining the hierarchical drill-down semantics and the `note` on empty results. It also clarifies that `contains` filters a path substring. This is useful behavior beyond the structured fields.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Three sentences, all dense with useful information, and the core purpose is front-loaded. The use case and parameter examples are not filler; they directly help invocation correctness.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given the output schema exists and annotations already carry safety and idempotency, this description covers the remaining invocation-relevant context: topic choices, hierarchy levels, contains behavior, and empty-result handling. Nothing critical is missing for this read-only drill-down tool.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters5/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 0%, so the description carries the full burden for parameter meaning. It explains topic with concrete examples, level with hierarchy depth semantics, and contains with a path-substring filter. Year is at least mentioned as part of the core use case, which compensates for the bare schema.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description opens with a specific noun phrase: 'Hierarchical federal-budget breakdown for one topic and year', which names the resource and the operation. It also gives a concrete use case and topic examples, so an agent can identify what this tool returns and how it differs from sibling tools like fiscal_by_institution or fiscal_headline.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description explains when to use it: 'see where the money goes', with a concrete example question and sample topic values. It does not explicitly name alternatives or state when not to use it, but the context is clear enough that a capable agent can route correctly.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

fiscal_by_institutionFiscal By InstitutionA
Read-onlyIdempotent

Federal spending by department / administrative unit since 2007.

Use case: compare personnel, IT or external-services spending across departments β€” e.g. "IT spending of the Finanzdepartement since 2010?". variable one of: 'Personalausgaben', 'Informatik', 'Beratung und externe Dienstleistungen', 'Anzahl Vollzeitstellen'. An empty result carries a note with guidance.

ParametersJSON Schema
NameRequiredDescriptionDefault
year_toNo
variableNoPersonalausgaben
year_fromNo
departementNo

Output Schema

ParametersJSON Schema
NameRequiredDescription
noteNoguidance when the result is empty or has a caveat (ARCH-003)
pointsYes
sourceNo
provenanceYesdump = freshly fetched CSV, cached = in-memory
filter_variableYes
filter_departementYes

TDQS

A4.1/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations already declare the tool read-only and idempotent, so the description does not need to repeat that. It adds meaningful behavior beyond annotations: the data starts in 2007, and empty results carry a `note` with guidance, which is valuable operational context for an agent.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is compact and every sentence earns its place: definition, use case, parameter guidance, and empty-result handling. It is front-loaded with the core purpose and contains no filler or redundancy.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a read-only query tool with a rich annotation set and an output schema, this is nearly complete: it gives valid variable values, data range, an illustrative query, and empty-result behavior. The main gap is not mentioning how to discover valid `departement` values or when to prefer sibling tools like `fiscal_list_dimensions`.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters4/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 0%, so the description must compensate. It enumerates the accepted values for `variable`, and the example 'IT spending of the Finanzdepartement since 2010?' gives practical meaning to `departement` and the year range parameters. It does not specify valid `departement` spellings, but it still adds substantial semantic value beyond the bare schema.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description identifies the resource (federal spending data) and the unit of analysis (department/administrative unit), and the use case clarifies it is for cross-department comparisons. However, it lacks an explicit action verb like 'returns' or 'lists' and never names sibling tools to differentiate them, so it stops short of a 5.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The 'Use case' line explicitly states when this tool is appropriate β€” comparing personnel, IT, or external-services spending across departments β€” and gives a concrete query example. It does not state exclusions or point to alternative sibling tools, so it misses the explicit when-not-to-use guidance that would earn a 5.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

fiscal_headlineFiscal HeadlineA
Read-onlyIdempotent

Headline fiscal time series: revenue, expenditure, balance and debt ratios from 1990 to the latest year the EFV publishes, actuals and forward-looking years alike. Read is_projection per point to tell them apart; not every household carries forward years.

Use case: track how a federal aggregate evolved over time β€” e.g. "how did the Bund balance develop since the 2022 rate turnaround?". variable e.g. 'saldo', 'einnahmen', 'ausgaben', 'bruttoschuldenquote'. household: bund|ktn|gdn|staat|sv. model: fs|gfs. Every point flags is_projection. Call fiscal_list_dimensions first to discover valid values; an empty result carries a note with guidance.

ParametersJSON Schema
NameRequiredDescriptionDefault
modelNofs
year_toNo
variableYes
householdNobund
year_fromNo

Output Schema

ParametersJSON Schema
NameRequiredDescription
noteNoguidance when the result is empty or has a caveat (ARCH-003)
unitNo
modelYesfs (Finanzstatistik) | gfs (GFS-Modell)
pointsYes
sourceNo
variableYes
householdYeshh: bund | ktn | gdn | staat | sv | bund_ktn_gdn
provenanceYesdump = freshly fetched CSV, cached = in-memory

TDQS

A3.9/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations already declare readOnly, idempotent, and openWorld hints. The description adds meaningful behavioral context: it includes both actuals and forward-looking years, each point carries an is_projection flag, and not every household has forward years. This goes beyond the annotations and helps the agent interpret results.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is two paragraphs with a clear lead sentence, a use case, and parameter examples. It is not overly verbose and front-loads the core purpose. It could be slightly more compact but is well-structured overall.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

With an output schema present, the description doesn't need to detail return values beyond the is_projection flag it mentions. It covers the main parameters, points to fiscal_list_dimensions for valid values, and explains the data scope. The only minor gap is the incomplete description of year_from/year_to, but the overall guidance is sufficient for correct invocation.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 0%, so the description must compensate. It provides example values for variable, household, and model (e.g., 'saldo', 'bund', 'fs'), which helps. However, it does not explicitly explain year_from and year_to, their defaults, or the meaning of the null values. This leaves gaps for two of the five parameters.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly states the tool returns fiscal time series (revenue, expenditure, balance, debt ratios) from 1990 onward, with a concrete use case. It does not explicitly differentiate from sibling fiscal tools like fiscal_budget_breakdown or fiscal_by_institution, so it earns a 4 rather than a 5.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description provides a specific use case ('track how a federal aggregate evolved over time') and instructs the agent to call fiscal_list_dimensions first to discover valid values. It does not list when not to use this tool or name alternatives, so it falls short of a 5 but is still clear context.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

fiscal_list_dimensionsFiscal List DimensionsA
Read-onlyIdempotent

List the valid dimension values across all datasets (variables, households, models, budget topics, departments).

Use case: call this first to build correct parameters for the other tools β€” it turns free-text guesses into exact filter values. Loads all three dumps, so it may take a moment on a cold cache.

ParametersJSON Schema
NameRequiredDescriptionDefault

No parameters

Output Schema

ParametersJSON Schema
NameRequiredDescription
modelsYes
sourceNo
householdsYes
provenanceYesdump = freshly fetched CSV, cached = in-memory
budget_topicsYes
headline_variablesYes
institution_variablesYes
institution_departmentsYes

TDQS

A4.5/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations already declare readOnlyHint, openWorldHint, idempotentHint, and non-destructiveness. The description adds the useful behavioral caveat that it loads all three dumps and may be slow on a cold cache, which goes beyond the annotations without contradicting them.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Two short, front-loaded sentences: the first states the core purpose and scope, the second gives the use case and a performance caveat. Every sentence adds value with no redundancy.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a zero-parameter discovery tool with an output schema, the description covers purpose, scope, use case, and a relevant performance cost. Nothing essential is missing for an agent to select and call it correctly.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters4/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

The tool takes zero parameters, so the schema has nothing to document and the description need not compensate. The description still reinforces its role by noting it produces exact filter values that feed parameters into other tools.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description names a specific action ('List') and a clear resource ('valid dimension values across all datasets'), enumerating the covered categories. It also distinguishes itself from sibling query tools by positioning this as the discovery step that builds correct parameters.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description explicitly says when to call it first: to turn free-text guesses into exact filter values for other tools. It gives clear context for its use, though it does not explicitly name alternatives or state when not to use it.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

fiscal_statusFiscal StatusA
Read-onlyIdempotent

Report cache freshness and upstream health per dataset.

Use case: check whether the data is fresh, cached or degraded before trusting a figure β€” the health endpoint of this server. Never returns empty silently; used for graceful degradation.

ParametersJSON Schema
NameRequiredDescriptionDefault

No parameters

Output Schema

ParametersJSON Schema
NameRequiredDescription
sourceNo
healthyYes
messageYes
datasetsYes

TDQS

A4.5/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations already cover read-only, non-destructive, idempotent behavior. The description adds meaningful context beyond that: 'Never returns empty silently; used for graceful degradation,' which is a behavioral guarantee not present in the annotations. This goes beyond the baseline.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is extremely concise: two sentences, with the core purpose front-loaded in the first line and the use case and behavior in the second. Every sentence adds value and there is no filler.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a zero-parameter health endpoint with a rich output schema and strong annotations, the description covers purpose, usage, and behavioral nuance. There are no missing pieces an agent would need to call it correctly.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters4/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

The tool takes zero parameters, so the description is not required to explain parameter behavior. The baseline of 4 applies, and the description adds no parameter-related information but also does not need to.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description states a specific verb and resource: 'Report cache freshness and upstream health per dataset.' It also explicitly labels itself as 'the health endpoint of this server,' distinguishing it from the sibling data-querying tools like fiscal_headline and fiscal_budget_breakdown.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description gives a clear use case: 'check whether the data is fresh, cached or degraded before trusting a figure.' This tells the agent when to use it, though it does not explicitly mention alternatives or when not to use it, so it stops short of a full 5.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Tool Schema Changelog

Recent tool additions, removals, and schema changes observed during successful MCP inspections.

  1. 1 tool updatev0.4.0
    • Changedfiscal_headline1 field changed
      • changedOutput schema / properties / points / items / properties / kind / description
        Previous value: -"raw EFV source label, e.g. 'Financial statements', 'Budget/financial plans'"New value: +"raw EFV source label, passed through verbatim β€” e.g. 'Rechnung', 'Prognosen'. The source picks its own wording and has switched language before (English until 2026-08-27), so branch on `is_projection` rather than on this string."
  2. 6 tool updatesv0.3.1
    • First observeddump_status
    • First observedfiscal_budget_breakdown
    • First observedfiscal_by_institution
    • First observedfiscal_headline
    • First observedfiscal_list_dimensions
    • First observedfiscal_status

TDQS

A4.4/5.0

Scored across 6 tools

Disambiguation5/5

Each tool targets a distinct aspect of the fiscal data: time series, budget breakdown, institutional spending, dimension discovery, and health status. The only potential overlap (fiscal_status vs. dump_status) is explicitly resolved by deprecating the latter, so agents can clearly distinguish them.

Naming Consistency5/5

All tools follow the `fiscal_` prefix with a descriptive noun (headline, budget_breakdown, by_institution, list_dimensions, status). The deprecated alias `dump_status` breaks the pattern but is clearly marked as temporary, so the overall convention remains highly consistent.

Tool Count5/5

Six tools cover the core operations for a read-only fiscal data server: querying aggregates, hierarchical breakdowns, institutional comparisons, dimension discovery, and health checks. This is well-scoped without unnecessary bloat or gaps.

Completeness5/5

The surface covers the primary data access patterns for the domain: time-series aggregates, budget hierarchies, departmental spending, and metadata discovery. The status tool ensures graceful degradation, and the deprecated alias is a minor artifact that does not affect completeness.

Maintenance

ActivityActive
ResponsivenessResponsive

Related MCP Connectors

Related MCP Servers

  • A
    license
    A
    quality
    B
    maintenance
    MCP server for Swiss federal legislation metadata via Fedlex, enabling search and retrieval of act details with ELI URIs, SR numbers, and multilingual support.
    3
    42 PyPI
    Apache 2.0
  • F
    license
    A
    quality
    C
    maintenance
    MCP server for the Swiss federal tax calculator, providing income, wealth, inheritance, and corporate tax figures for all Swiss municipalities and tax years 2010-2026.
    13
    2
    -
  • A
    license
    A
    quality
    A
    maintenance
    MCP server for Swiss electricity data from three official sources β€” production mix, consumption forecast, storage-lake fill, consumer price index, tariffs per municipality, and dataset discovery. Zero authentication.
    12
    136 PyPI
    MIT