docx_to_markdown
Convert DOCX documents to Markdown with pipe tables and strikethrough, then check output fidelity and flag content Markdown cannot carry.
Instructions
Markdown (CommonMark + pipe tables + strikethrough) from a DOCX, via sandboxed Pandoc. The report checks text, numbers, identifiers, links, headings, lists, tables and bold/italic, and lists what Markdown cannot carry (headers/footers, comments, equations, images). [docbridge schema 0.2.4]
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| detail | No | summary: report_summary + report_id (get_report has the rest). full: the whole report inline. | summary |
| overwrite | No | ||
| input_path | Yes | Absolute file path. | |
| output_path | No | ||
| report_path | No | ||
| max_differences | No | Differences kept in the full report; the rest are counted. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| tool | Yes | ||
| error | No | ||
| report | No | The full report. Over MCP only with detail='full'. | |
| outputs | No | ||
| report_id | No | Pass to get_report for the full report (held for this server session). | |
| disclaimer | No | docbridge never alters, summarizes or silently truncates source evidence. docbridge reports what it compared. PASS on an axis covers only that axis's stated scope; NOT_CHECKED and UNSUPPORTED are never passes. PDF text is the text layer as PyMuPDF decodes it, not the rendered glyphs. Nothing here interprets meaning: numbers are compared as characters, not as values. | |
| report_path | No | ||
| report_summary | No | ||
| schema_version | Yes | ||
| conversion_notes | No | ||
| operation_completed | Yes | The operation ran and wrote its outputs. Says NOTHING about fidelity: read the report status. |