kubik-tools
Server Details
Freight calculators (weight, metres, vehicle fit) and authenticated team packing-library tools.
- Status
- Healthy
- Last Tested
- Transport
- Streamable HTTP
- URL
Glama MCP Gateway
Connect through Glama MCP Gateway for full control over tool access and complete visibility into every call.
Full call logging
Every tool call is logged with complete inputs and outputs, so you can debug issues and audit what your agents are doing.
Tool access control
Enable or disable individual tools per connector, so you decide what your agents can and cannot do.
Managed credentials
Glama handles OAuth flows, token storage, and automatic rotation, so credentials never expire on your clients.
Usage analytics
See which tools your agents call, how often, and when, so you can understand usage patterns and catch anomalies.
Tool Definition Quality
Average 4.8/5 across 7 of 7 tools scored.
Each tool targets a distinct resource and action: chargeable weight, loading metres, vehicle fit, article creation, observation logging, quantity resolution, and undo. Even the similar freight calculations are clearly differentiated by their output (weight vs floor space vs fit check) with explicit cross-references.
All tool names follow a consistent verb_noun pattern (calculate_, check_, create_, log_, resolve_, undo_) using snake_case throughout. The verbs accurately reflect the operation and the nouns identify the domain object, making the set predictable.
Seven tools is well-scoped for the server's dual purpose of freight calculations and article packing management. Each tool earns its place—three for freight math, three for article data, and one for undo support—with no redundancy.
The freight calculation side is complete for common scenarios (weight, LDM, vehicle fit), and the article side covers creation, observation logging, and resolution. Minor gaps exist—no direct article list/update/delete beyond undoing a recent create, and no raw observation history—but these are partially addressed via the app and the undo window.
Available Tools
7 toolscalculate_chargeable_weightCalculate Chargeable WeightARead-onlyIdempotentInspect
Calculates chargeable (billable) weight -- frachtpflichtiges Gewicht, Frachtgewicht -- for a freight shipment (Stückgut or Sammelgut) from a list of cargo pieces, for one of four transport modes. Answers questions like "wie viel wiegt die Sendung frachtpflichtig" or "was ist das Volumengewicht".
For each mode, chargeable weight is the greater of the actual (scale) weight and the volumetric weight (Volumengewicht), where volumetric weight is derived from total volume using a mode-specific default divisor (overridable via volumetric_divisor):
air (Luftfracht): volume_cm3 / 6000 (IATA standard, 167 kg/m3)
courier: volume_cm3 / 5000 (common express-carrier convention, e.g. DHL/FedEx/UPS)
road: volume_m3 * 333 (simple volumetric "1:3" convention; does not model Lademeter/LDM-based road pricing -- for loading-metre, Stellplätze, or vehicle-fit questions, use calculate_loading_metres and check_truck_fit instead, both on this server)
sea_lcl (Seefracht): volume_m3 * 1000 (W/M -- weight or measurement, 1 revenue tonne per m3)
Worked example: 2 pieces, 60x40x50cm, 45 kg each, air mode -> total actual weight 90 kg, total volume 0.24 m3, volumetric weight 40 kg (240,000 cm3 / 6000) -> chargeable weight 90 kg (actual weight governs, since it exceeds the volumetric weight).
Rounding: air and courier chargeable/volumetric weight round UP to the nearest 0.5 kg (chargeable_weight_raw_kg gives the unrounded value, chargeable_weight_kg the rounded one). Road and sea_lcl are not rounded up, just reported to 1 decimal place.
Edge cases: missing or invalid mode, more than 100 pieces, or any non-positive dimension/weight/quantity returns a clear, structured explanation rather than an error stack -- never a guessed default mode or divisor.
Returns total actual weight, total volume, volumetric weight, raw and rounded chargeable weight, which one governs, the divisor used, and a one-line human-readable summary.
| Name | Required | Description | Default |
|---|---|---|---|
| mode | No | Transport mode (required): one of "air" (Luftfracht), "road", "sea_lcl" (Seefracht), "courier". Selects the default volumetric divisor: air=6000 cm³/kg (IATA), courier=5000 cm³/kg, road=333 kg/m³ (simple volumetric convention — for loading-metre/Lademeter/LDM and vehicle-fit questions, use calculate_loading_metres and check_truck_fit instead, both on this server), sea_lcl=1000 kg/m³ (W/M). | |
| pieces | No | List of cargo pieces (up to 100 line items, required, non-empty). Use quantity to combine identical pieces rather than repeating rows. | |
| volumetric_divisor | No | Optional positive override of the mode's default divisor. For air/courier this is cm³ per kg (divided into volume); for road/sea_lcl this is kg per m³ (multiplied by volume). |
Tool Definition Quality
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotations declare readOnlyHint=true and destructiveHint=false, so the agent knows it's a safe read operation. The description adds substantial behavioral context: rounding rules (UP to nearest 0.5 kg for air/courier, no rounding for road/sea_lcl), edge-case handling (returns structured explanation, never guesses defaults), and the output object's contents.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is long but every sentence earns its place. It is structured with clear sections: formula, per-mode divisors, a worked example, rounding, edge cases, and returns — making it easy to scan while being exhaustive.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
With no output schema, the description fully enumerates return fields (total actual weight, total volume, volumetric weight, raw/rounded chargeable weight, governing dimension, divisor, summary). It also covers edge cases and mode-specific behavior, making the tool safe and predictable to invoke.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
The input schema already has 100% coverage, but the description adds meaning: explains how mode selects divisors, how volumetric_divisor overrides defaults, and provides the actual formulas (e.g., volume_cm3 / 6000). It also explains the distinction between chargeable_weight_raw_kg and chargeable_weight_kg.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states a specific verb ('Calculates chargeable (billable) weight') with a defined resource (freight shipment from cargo pieces across four transport modes). It also differentiates from sibling tools by explicitly directing loading-metre questions to calculate_loading_metres/check_truck_fit.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description explicitly states when to use this tool for chargeable weight calculations and provides direct alternatives for loading-metre/vehicle-fit scenarios: 'for loading-metre, Stellplätze, or vehicle-fit questions, use calculate_loading_metres and check_truck_fit instead.' It also names the four applicable modes.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
calculate_loading_metresCalculate Loading MetresARead-onlyIdempotentInspect
Calculates loading metres (Lademeter, LDM) -- the floor-space unit (Ladefläche) that governs European road freight -- for a list of cargo pieces, plus the equivalent number of pallet places (Palettenstellplätze, Stellplätze). Answers questions like "wie viele Lademeter" or "wieviel Platz brauche ich im Lkw".
Units: length_cm and width_cm (Länge/Breite in cm) per piece. weight_kg_per_piece (Gewicht in kg) is optional. Formula: LDM = sum(length_cm x width_cm x quantity) / reference_deck_width_cm / 100. Height is irrelevant here -- LDM is a floor-footprint metric, not a volume one (for volumetric/chargeable weight, use calculate_chargeable_weight instead).
Reference width: Lademeter (LDM) is always computed at the fixed 240cm (2.4m) reference lane -- this is a commercial road-freight convention, not a DIN/EN/ISO/VDI standard, and does not vary by vehicle. If vehicle_id names one of the loadable vehicle profiles (e.g. "semi_136" = Sattelzug 13,6 m, "rigid_75" = 7,5-Tonner), the response additionally reports deck_length_m: how much of that specific vehicle's own deck length the cargo occupies -- a different, vehicle-specific quantity, not the LDM figure. Pallet-place equivalents are also computed at the 240cm reference lane.
Worked example: 8 Europaletten/EUR-Paletten (120cm x 80cm) at the standard 240cm width = (120808)/240/100 = 3.2 LDM, equivalent to 8 Palettenstellplätze (one Europalette occupies 0.4 LDM at this width). An Industriepalette (120x100cm) occupies more floor space per unit: 0.5 LDM at the same width.
Edge cases: an unknown vehicle_id returns the list of available vehicles instead of guessing -- it never silently picks one. If any piece omits weight_kg_per_piece, total_weight_kg is returned as null rather than an incomplete partial sum. Maximum 100 piece lines (use quantity to combine identical pieces).
| Name | Required | Description | Default |
|---|---|---|---|
| pieces | No | List of cargo pieces (up to 100 line items, required, non-empty). Use quantity to combine identical pieces rather than repeating rows. | |
| vehicle_id | No | Optional vehicle profile id (e.g. "semi_136" = Sattelzug 13,6 m, "rigid_75" = 7,5-Tonner). Does not change the LDM calculation (always the fixed 240cm reference lane) -- adds deck_length_m, how much of that vehicle's own deck length the cargo occupies. If given but not recognized, the response lists every available vehicle instead of guessing. |
Tool Definition Quality
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Even with annotations (readOnlyHint, idempotentHint, destructiveHint), the description adds rich behavioral context: unknown vehicle_id returns a list instead of guessing, missing weight yields null for total_weight_kg, max 100 piece lines, and the fixed 240cm reference width. No contradictions.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is well-structured and front-loaded with a clear opening sentence, but it is quite long and repeats some vehicle_id behavior already present in the schema. Every sentence adds value, though the length could be slightly reduced without losing information.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the absence of an output schema, the description thoroughly covers all relevant return values (LDM, pallet places, deck_length_m, total_weight_kg) and edge cases. It also includes the formula, units, and a worked example, making it complete for an agent to invoke correctly.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema coverage is 100%, but the description adds meaningful semantics beyond it: the formula, worked examples, and optionality/behavior of weight and vehicle_id. While the schema already documents the fields, the description reinforces how values are used in the calculation.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states the tool calculates loading metres (LDM) and pallet-place equivalents, with specific verb, resource, and scope. It also explicitly differentiates from calculate_chargeable_weight, making it distinct from siblings.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
Provides explicit guidance on when to use the tool (floor-space questions, not volumetric weight) and names the alternative tool for volumetric/chargeable weight. Also explains when vehicle_id is relevant (to get deck_length_m).
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
check_truck_fitCheck Truck FitARead-onlyIdempotentInspect
Checks whether a list of cargo pieces fits a named vehicle by loading metres (Lademeter, LDM, Stellplätze) and, if weight is given, by payload (Zuladung, Nutzlast) too -- the two capacity limits that actually govern road freight ("passt das auf..."), not full 3D placement or Ladungssicherung (load securing).
Vehicle profiles (vehicle_id -> German name): "semi_136" = Sattelzug/Sattelauflieger 13,6 m (colloquially also "40-Tonner"), "curtain_136" = Planensattel 13,6 m, "flatbed_136" = Pritsche 13,6 m, "rigid_75" = 7,5-Tonner, "midi_12" = 12-Tonner/Koffer-Lkw, "rigid_18" = 18-Tonner, "rigid_26" = 26-Tonner (3-Achser), "drawbar_40" = Hängerzug, "sprinter_l3h2" = Mercedes Sprinter L3H2 (3,5t), "ducato_l4h2" = Fiat Ducato L4H2 (3,5t), "cont_20"/"cont_40"/"cont_40hc" = 20-/40-/40-Fuß-HC-Container. Cargo like a Gitterbox/Rollbehälter is just another piece by footprint -- no separate cargo-type parameter needed.
Units: length_cm and width_cm per piece in centimetres (Länge/Breite in cm), weight_kg_per_piece in kilograms (Gewicht in kg, optional -- if omitted, only the floor-space/LDM check runs; the payload check is honestly skipped, not guessed).
Three modes, by what's given: (1) vehicle_id + pieces -> full fit check; if it does NOT fit, the response additionally includes recommended_vehicles -- the smallest fitting alternatives by payload, e.g. cargo that overloads a 7,5-Tonner might fit a 12-Tonner or 18-Tonner instead (suggest, never auto-pick). (2) vehicle_id alone, no pieces -> that vehicle's own payload_kg/max_ldm_m/deck_width_cm, a plain spec lookup (e.g. "wie viele Stellplätze hat ein Standard-Sattelzug" or "maximale Zuladung Sattelzug"). (3) neither given, or pieces given with no vehicle_id -> the full vehicle list, filtered to fitting ones when pieces were given. An unrecognized vehicle_id also returns the full list -- it never assumes which vehicle you mean.
Worked example: 6 Europaletten (120x80cm), 4.8 tonnes total, against vehicle_id "rigid_75" -- loading metres (2.5 LDM) are well within the 6.2m limit, but 4800kg exceeds the 3500kg payload, so fits=false, limiting_constraint="payload", recommended_vehicles lists "midi_12" (12-Tonner) and "rigid_18" (18-Tonner) first -- the smallest vehicles that hold 4.8t.
Edge cases: if no vehicle profile fits the given cargo at all, says so honestly ("exceeds all vehicle profiles") rather than recommending an impossible option. Utilisation percentages are uncapped on purpose (an overloaded plan reads "137%", not a reassuring clamped "100%"). Maximum 100 piece lines.
| Name | Required | Description | Default |
|---|---|---|---|
| pieces | No | List of cargo pieces (up to 100 line items), optionally omitted entirely to just look up a vehicle's specs (with vehicle_id) or list all vehicles (without). Use quantity to combine identical pieces rather than repeating rows. | |
| vehicle_id | No | Vehicle profile id to check against or look up (e.g. "semi_136" = Sattelzug 13,6 m, "rigid_75" = 7,5-Tonner). If omitted or not recognized, the response lists every available vehicle instead of guessing. |
Tool Definition Quality
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Even though annotations already declare readOnly and idempotent hints, the description adds substantial behavioral context: it explains that payload check is honestly skipped if weight is omitted, that recommended_vehicles are suggestions never auto-picked, that utilization percentages are uncapped intentionally, and that invalid vehicle IDs return the full list rather than guessing. No contradictions with annotations.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is long but well-structured with clear sections: intro, vehicle profiles, units, modes, example, and edge cases. The first sentence immediately states the core purpose, and each section earns its place given the tool's complexity. It is slightly verbose, but no sentence is wasted; a 4 is appropriate for this density of necessary information.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Despite having only 2 parameters and no output schema, the description covers all necessary context: input modes, units, vehicle profile mapping, edge cases (unrecognized ID, no fitting vehicle), limits (100 pieces), and even an example. The tool's behavior across all possible input combinations is fully documented, making it complete for an agent to invoke correctly.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema coverage is 100%, giving the baseline 3, but the description significantly enriches both parameters. It explains the meaning of pieces as lines with quantity, provides units (cm, kg), clarifies weight_kg_per_piece as optional, and details how vehicle_id maps to German names and behavior when omitted. It also provides a worked example that demonstrates parameter usage, far exceeding schema descriptions.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states the tool checks cargo fit against vehicle capacity limits (loading metres and payload), with a specific verb ('checks') and resource ('list of cargo pieces fits a named vehicle'). It explicitly distinguishes itself from full 3D placement or load securing, and from sibling tools like calculate_loading_metres. The purpose is unmistakable and specific.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description provides explicit usage guidance by defining three modes based on input combinations: vehicle_id + pieces for full check, vehicle_id alone for spec lookup, and no vehicle_id for listing vehicles. It also tells users when this tool is appropriate versus alternatives by excluding 3D placement and load securing, and by referencing sibling tools implicitly through the mention of loading metres. This is exemplary guidance.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
create_article_profileCreate Article ProfileAInspect
Creates a new article in your team's kubik.tools Library -- team-scoped from creation, visible to the whole team immediately (not a personal draft).
Two-step by design, never a silent write: call once WITHOUT confirm to get a preview of exactly what would be created (including a check for an existing article that looks like a likely duplicate); call again with confirm=true and the preview_id from the first response to actually create it. A confirmed creation can be undone within 15 minutes via undo_change -- after that, edit or delete it directly in the app.
On the confirm call, only api_key, confirm, and preview_id are actually read -- every other field is required/optional per the schema for shape-consistency but ignored if resupplied, since the values captured during the preview call are what gets created; to change any of them, call again without confirm for a fresh preview.
Do not use this to log a real packing observation for an article that already exists -- use log_pack_observation instead; this tool only creates the article record itself, never packing data. Do not skip the unconfirmed preview call even if you are confident there is no duplicate -- the duplicate check only runs on that first call, and confirm=true without a fresh preview_id will be rejected.
Required: name, supplier. Optional: article_number, hs_code (exactly 8 digits if given), description_de, description_en.
| Name | Required | Description | Default |
|---|---|---|---|
| name | Yes | Article name, e.g. 'Rotor hub casting'. Required. | |
| api_key | Yes | Your team's kubik.tools MCP API key (kubik_mcp_...). Required. | |
| confirm | No | Set true, with preview_id, to actually create the article after reviewing the preview. Omit or false for a dry-run preview only. | |
| hs_code | No | HS/tariff code, exactly 8 digits. Format-only check -- not verified against a real tariff database. | |
| supplier | Yes | Supplier name. Required. | |
| preview_id | No | The preview_id returned by the first (unconfirmed) call. Required when confirm=true. | |
| article_number | No | Supplier's own article/SKU number, if known. | |
| description_de | No | ||
| description_en | No |
Tool Definition Quality
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotations only provide negative hints (not readOnly, not idempotent, not destructive), so description carries full burden. It discloses the two-step workflow, that the tool is never a silent write, that parameters are ignored on the confirm call except api_key/confirm/preview_id, that the duplicate check only runs on the first call, and that creation is team-visible immediately. This is rich, non-obvious behavior.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is long but well-organized, front-loading the core purpose and then explaining the two-step design, parameter behavior, and exclusions. Each paragraph earns its place, though the final required/optional list somewhat duplicates schema information. Slightly verbose but justified by the tool's complexity.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Despite having no output schema, the description conveys the preview/confirm response flow, duplicate-check behavior, undo timeframe, and team-scoped visibility. It fully addresses the tool's complexity, including the subtle behavior that only preview_id and confirm matter on the second call, leaving no ambiguity about how to invoke it correctly.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema covers most parameter descriptions, but the description adds critical interplay semantics: only api_key, confirm, and preview_id are read on the confirm call, all other fields are ignored if resupplied, and preview_id is required when confirm=true. It also clarifies hs_code must be exactly 8 digits and reiterates required/optional fields. This goes well beyond schema descriptions.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description opens with 'Creates a new article in your team's kubik.tools Library', using a specific verb and resource, and immediately clarifies team-scoping. It distinguishes itself from sibling tools by explicitly stating it only creates article records, never packing data (unlike log_pack_observation).
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
Provides explicit when-to-use guidance: 'Do not use this to log a real packing observation... use log_pack_observation instead'. Also explains the mandatory two-step preview/confirm flow and warns against skipping the unconfirmed call, plus mentions undo_change for reversal within 15 minutes.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
log_pack_observationLog Pack ObservationAInspect
Logs a real, observed consolidation pack -- "these articles, at these quantities, actually packed onto N pallets at these dims/weight" -- against an existing consolidation group in your team's Library. This is ground truth: it becomes a new data point resolve_quantities and the app's own resolve flow learn from (PRINCIPLES.md P-13).
Two-step by design: call once WITHOUT confirm to preview exactly what would be logged; call again with confirm=true and the preview_id to actually log it. A confirmed log can be undone within 15 minutes via undo_change -- after that it's permanent (append-only ground truth, by design -- see PRINCIPLES.md P-17/P-18 for why).
On the confirm call, only api_key, confirm, and preview_id are actually read -- every other field is required by the schema for shape-consistency but ignored if resupplied, since the values captured during the preview call are what gets logged; to change any of them, call again without confirm for a fresh preview.
Do not use this for a single article's own packing history outside a consolidation group -- that data comes from the app's own data entry, not this tool. Do not use this to correct a mistaken past observation after the 15-minute undo window -- log a new, correct observation instead; past ones are never edited. Requires an EXISTING consolidation_groups id -- this tool cannot create a new consolidation group.
Requires an existing consolidation_groups id (from the app's Consolidation Groups screen) and each member article's number (resolved to its profile automatically).
| Name | Required | Description | Default |
|---|---|---|---|
| api_key | Yes | Your team's kubik.tools MCP API key (kubik_mcp_...). Required. | |
| confirm | No | Set true, with preview_id, to actually log the observation after reviewing the preview. Omit or false for a dry-run preview only. | |
| width_cm | Yes | ||
| height_cm | Yes | ||
| length_cm | Yes | ||
| weight_kg | Yes | Total gross weight, in kg. | |
| preview_id | No | The preview_id returned by the first (unconfirmed) call. Required when confirm=true. | |
| pallet_count | Yes | Number of pallets this pack used. | |
| member_quantities | Yes | The articles and quantities actually packed together, e.g. [{article_number: 'MLB1001', qty: 3}]. | |
| consolidation_group_id | Yes | The id of an existing consolidation group in your team's Library. Required. |
Tool Definition Quality
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotations are all false/neutral, so the description carries the full burden of behavioral disclosure. It richly describes the two-step design, the confirm-call field-ignoring behavior, the 15-minute undo window, append-only permanence, and the fact that it cannot create a new group. No contradictions with annotations.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is long but every sentence earns its place: it starts with the core function, then the workflow, then field semantics, exclusions, and prerequisites. It is well-structured and information-dense without waste, appropriate for a complex tool with 10 parameters and a two-step flow.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
For a tool with 10 parameters, no output schema, and non-trivial behavioral rules, the description covers purpose, flow, field roles, undo/persistence behavior, exclusions, and prerequisites. The only minor gap is the exact preview response shape, but the description indicates what the preview is for and returns preview_id, which is sufficient for the agent to use the tool correctly.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema coverage is 70%, and the description adds meaningful context by explaining which fields are read on confirm (api_key, confirm, preview_id), showing an example for member_quantities, and clarifying preview_id's role. The three dimension params (length_cm, width_cm, height_cm) lack individual descriptions, but their names and units are self-explanatory and the description refers to 'dims/weight', so the meaning is adequately conveyed.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description opens with a specific verb+resource: 'Logs a real, observed consolidation pack...against an existing consolidation group in your team's Library.' It clearly distinguishes this from sibling tools by stating it produces ground-truth data for resolve_quantities and explicitly says 'Do not use this for a single article's own packing history' and 'Do not use this to correct a mistaken past observation,' making the purpose unambiguous.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description gives explicit when-to-use context ('real, observed consolidation pack'), prerequisites (existing consolidation_groups id, article numbers), and when-not-to-use exclusions (single-article history, post-undo corrections). It also explains the two-step preview/confirm flow and references alternatives like undo_change, providing complete usage guidance.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
resolve_quantitiesResolve Article QuantitiesAInspect
Resolves an ordered quantity of a Library article into concrete packed units (boxes/pallets), using that article's own logged packing history -- never a generic guess.
Check resolution_status first: "resolved" means units/tier are a real answer, safe to use directly. "ambiguous" means the article has more than one viable packaging family (e.g. logged both as pallets and as crates) with meaningfully different pack counts -- units and tier are both null, and alternatives lists every viable option with no default among them. When resolution_status is "ambiguous", do not select an alternative autonomously -- present the options to the user and ask which packaging family they mean; do not guess based on which one appears first. "not_found" means the article has no packed-form data logged at all -- there is nothing to resolve, alternatives is absent, and units/tier are null. needs_confirmation (boolean) is kept only for backward compatibility with callers written before resolution_status existed -- new integrations should check resolution_status.
When resolution_status is "resolved", tier is one of: "observed" (an exact match against a real logged pack), "estimated" (interpolated between two real logged points), or "verify" (extrapolated beyond the highest -- or below the lowest -- quantity ever actually logged for this article; still returns a real number, but flagged as needing a human's eyes before it's trusted).
Does not resolve prepack-table or consolidation-group quantities -- those use the same tier engine but a different lookup key (a table/group id, not a free-text article search) and are not covered by this tool.
Requires an MCP API key (Authorization: Bearer ) issued for a kubik.tools team. Looks the article up by article number or name within that team's own Library -- never across teams.
| Name | Required | Description | Default |
|---|---|---|---|
| api_key | Yes | Your team's kubik.tools MCP API key (kubik_mcp_...). Required. | |
| quantity | Yes | The ordered quantity to resolve, in whole units of the article. Must be a positive number. | |
| article_query | Yes | The article number or name to search for in the Library, e.g. 'MLB1001' or 'rim holder'. |
Tool Definition Quality
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
The description explains in detail the resolution_status values ('resolved', 'ambiguous', 'not_found') and what each means, including how to handle 'ambiguous' (present options to the user, not guess). It also describes tier values ('observed', 'estimated', 'verify') with their implications and notes the backward-compatibility needs_confirmation field, providing rich behavioral context beyond the sparse annotations.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is long but well-structured, with clear sections for purpose, status handling, tiers, exclusions, and auth. Every sentence adds necessary behavioral detail, though the backward-compatibility note could be trimmed; it is not as tight as a two-sentence description but remains efficient for the complexity.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given there is no output schema and only three parameters, the description effectively explains all return semantics (units, tier, alternatives, needs_confirmation) and when they apply. It also covers exclusion cases and authentication, making it highly complete for a tool with this complexity.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema coverage is 100% for all three parameters, so the schema carries the load. The description adds minor context about the api_key (team-scoped Bearer token) and article_query (looks up by number or name within the team's Library), but largely repeats schema information, so it does not significantly raise the baseline.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states a specific verb and resource: 'Resolves an ordered quantity of a Library article into concrete packed units (boxes/pallets)' and explicitly distinguishes from a sibling concern by excluding prepack-table and consolidation-group quantities. It provides a concrete outcome and avoids any tautology.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description gives explicit guidance: check resolution_status first, do not autonomously choose when ambiguous, and clarifies that prepack/consolidation quantities use a different tool. It also states the API key requirement and team scoping, which are actionable usage prerequisites.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
undo_changeUndo Recent ChangeADestructiveInspect
Undoes a create_article_profile or log_pack_observation call, but only within 15 minutes of when it was confirmed, and only if nothing else now depends on it.
For article_profiles: refuses if any packed_forms, consolidation_group_members, or edit-history rows now reference the article (edit it or delete it manually in the app instead of undoing).
For consolidation_pack_observations: append-only ground truth past the 15-minute window, by design (PRINCIPLES.md P-17/P-18) -- undo only exists for a mistake caught immediately after logging it, never as a general edit/delete capability.
| Name | Required | Description | Default |
|---|---|---|---|
| table | Yes | Which kind of change to undo, matching the undo_token's origin (create_article_profile -> article_profiles, log_pack_observation -> consolidation_pack_observations). | |
| api_key | Yes | Your team's kubik.tools MCP API key (kubik_mcp_...). Required. | |
| undo_token | Yes | The undo_token returned by the original confirmed create_article_profile or log_pack_observation call. |
Tool Definition Quality
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Although annotations already mark destructiveHint=true, the description adds essential behavioral detail: refusal conditions for dependent article_profiles, the append-only nature of consolidation_pack_observations, and the design rationale referencing PRINCIPLES.md. No contradiction with annotations.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is appropriately sized for a tool with significant constraints. The first sentence front-loads the core action, and the subsequent paragraphs efficiently explain exceptions and design rationale without filler.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the tool is destructive, has no output schema, and handles two distinct table types, the description covers behavior, constraints, refusal cases, and rationale well. It doesn't describe the return value/success response, but the rest is complete enough for an agent to use safely.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
The input schema already describes all three parameters thoroughly (100% coverage), so the description adds limited extra parameter meaning. It does add a useful mapping between the 'table' enum and the originating call, but this is modest beyond the schema.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states it 'Undoes a create_article_profile or log_pack_observation call', naming specific verbs and resources. It also distinguishes this tool from its siblings by emphasizing that it is only for undo, not for general editing/deleting.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
It gives explicit usage boundaries: only within 15 minutes of confirmation, only if nothing depends on it, and never as a general edit/delete capability. It also names the alternative path: 'edit it or delete it manually in the app instead of undoing'.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
Claim this connector by publishing a /.well-known/glama.json file on your server's domain with the following structure:
{
"$schema": "https://glama.ai/mcp/schemas/connector.json",
"maintainers": [{ "email": "your-email@example.com" }]
}The email address must match the email associated with your Glama account. Once published, Glama will automatically detect and verify the file within a few minutes.
Control your server's listing on Glama, including description and metadata
Access analytics and receive server usage reports
Get monitoring and health status updates for your server
Feature your server to boost visibility and reach more users
For users:
Full audit trail – every tool call is logged with inputs and outputs for compliance and debugging
Granular tool control – enable or disable individual tools per connector to limit what your AI agents can do
Centralized credential management – store and rotate API keys and OAuth tokens in one place
Change alerts – get notified when a connector changes its schema, adds or removes tools, or updates tool definitions, so nothing breaks silently
For server owners:
Proven adoption – public usage metrics on your listing show real-world traction and build trust with prospective users
Tool-level analytics – see which tools are being used most, helping you prioritize development and documentation
Direct user feedback – users can report issues and suggest improvements through the listing, giving you a channel you would not have otherwise
The connector status is unhealthy when Glama is unable to successfully connect to the server. This can happen for several reasons:
The server is experiencing an outage
The URL of the server is wrong
Credentials required to access the server are missing or invalid
If you are the owner of this MCP connector and would like to make modifications to the listing, including providing test credentials for accessing the server, please contact support@glama.ai.
Discussions
No comments yet. Be the first to start the discussion!
Related MCP Servers
- Alicense-qualityCmaintenancePlan optimal container & truck loads: 3D layouts, right-size the container mix, and check utilization, centre of gravity, crush protection and securing across 200+ equipment types.17MIT
- AlicenseAqualityAmaintenanceAI agent access to 11 freight calculation and reference tools — LDM, CBM, chargeable weight, pallet fitting, ADR dangerous goods (2,939 entries), airline codes (6,352), HS codes (6,940), INCOTERMS, container specs, unit converter, and ADR 1.1.3.6 exemption calculator.241,3854MIT
- Flicense-qualityCmaintenanceReal published tariffs for European road freight and moving: quote by m3, kg, pallets or LDM across 560k+ routes. Dated price index, freight glossary, order submission. Live endpoint at https://mcp.fromtocargo.com/mcp (19 tools).
- AlicenseAqualityBmaintenanceEnables AI assistants to plan container and truck loads from plain-English shipment descriptions, returning fitted containers, utilization, non-fitting items, and interactive 3D load plans.195MIT