document_search
Search PDF, DOCX, MD, or TXT files for exact text or regex matches, returning verbatim results with page, line, offsets, context, and match IDs for follow-up reading.
Instructions
Exact search inside a PDF/DOCX/MD/TXT. Returns each match verbatim with its page/line/paragraph, character offsets, surrounding context and a match_id for document_read_around. Whitespace in the query matches any line break or spacing. Units that cannot be read are listed: no match there is not evidence of absence. [docbridge schema 0.2.4]
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| query | Yes | ||
| regex | No | ||
| encoding | No | Text encoding of a .md/.txt input. Omit for UTF-8; docbridge never guesses. | |
| input_path | Yes | Absolute file path. | |
| max_matches | No | ||
| start_match | No | ||
| context_chars | No | ||
| case_sensitive | No |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| tool | Yes | ||
| error | No | ||
| notes | No | ||
| source | No | ||
| matches | No | ||
| pattern | No | ||
| invariant | No | docbridge may reduce the volume of its own responses, but it never alters, summarizes, paraphrases, semantically filters, or silently truncates source evidence | |
| matches_total | No | ||
| case_sensitive | No | ||
| schema_version | Yes | ||
| matches_returned | No | ||
| next_start_match | No | ||
| unsearchable_units | No | ||
| operation_completed | Yes |