document_read
Read exact text from PDF, DOCX, MD, or TXT units by page, line, or paragraph range, returning offsets, a hash, and a cursor to resume whenever output is cut short.
Instructions
Verbatim text of units start..end (PDF pages, MD/TXT lines, DOCX paragraphs; end omitted = to the end), up to max_chars. Never summarized or filtered. The excerpt carries offsets and SHA-256; if cut, truncated=true, remaining_chars says how much is left and next_cursor continues exactly. [docbridge schema 0.2.4]
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| end | No | ||
| start | No | ||
| cursor | No | ||
| encoding | No | Text encoding of a .md/.txt input. Omit for UTF-8; docbridge never guesses. | |
| max_chars | No | Most characters returned; a cut is always stated, with a cursor to continue. | |
| input_path | Yes | Absolute file path. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| tool | Yes | ||
| error | No | ||
| match | No | ||
| source | No | ||
| excerpt | No | ||
| invariant | No | docbridge may reduce the volume of its own responses, but it never alters, summarizes, paraphrases, semantically filters, or silently truncates source evidence | |
| truncated | No | True if anything in the requested range was not returned. | |
| next_cursor | No | Pass as `cursor` to document_read to continue exactly. | |
| before_cursor | No | ||
| schema_version | Yes | ||
| remaining_chars | No | Characters of the requested range after the excerpt. | |
| requested_units | No | ||
| unreadable_units | No | Unreadable units inside this excerpt. | |
| operation_completed | Yes | ||
| omitted_before_chars | No | ||
| range_unreadable_units | No | Every unreadable unit of the whole requested range, returned or not. |